Skip to main content
DCS Global
Knowledge Base

Enterprise Infrastructure Glossary

Authoritative definitions for 80+ terms across data centers, AI infrastructure, critical power, networking, cybersecurity, and cloud. Written and maintained by DCS Global's engineering team.

80+Terms
10Categories
July 2026Updated
Engineer-VerifiedQuality

Showing 71 of 71 terms

2
1 term

2N Redundancy

Power Systems

A redundancy configuration where two complete, independent systems are available, each capable of supporting the full load. The highest level of redundancy for mission-critical infrastructure.

Uptime Institute Tier Standard
Related:N+1 RedundancyTier IVFault Tolerant
A
5 terms

AHU (Air Handling Unit)

Data Center

Mechanical unit that conditions and circulates air within a data center. Larger than CRAC units, typically connected to chilled water systems. Used in larger facilities where centralized cooling is more efficient.

Related:CRACCRAHChilled Water

All-Flash Array (AFA)

Storage

A storage system using only NAND flash (SSD) storage, providing dramatically higher IOPS and lower latency than hybrid or spinning disk arrays.

Related:NVMeIOPSLatency

ASHRAE TC 9.9

Standards

ASHRAE's Technical Committee 9.9 publishes thermal guidelines for data center equipment. Defines A1–A4 equipment classes with operating temperature and humidity ranges.

ASHRAE TC 9.9
Related:PUECoolingThermal Management

ATS (Automatic Transfer Switch)

Power Systems

A device that automatically transfers electrical load from a primary power source to a backup source (generator or alternate utility feed) upon detecting a failure. Transfer time: 4ms (static) to 30ms (mechanical).

NFPA 110 / IEEE C37.96
Related:UPSGeneratorPDU

Availability Zone

Data Center

An isolated location within a cloud region or data center campus with independent power, cooling, and network infrastructure. Designed to be failure-independent from other zones.

Related:RegionFault Domain
B
2 terms

BGP (Border Gateway Protocol)

Networking

The routing protocol of the internet. Used in data centers for external connectivity and increasingly for internal routing in large-scale spine-leaf fabrics.

RFC 4271
Related:OSPFECMPSpine-Leaf

Blanking Panel

Data Center

A 1U or 2U filler panel installed in empty rack spaces to prevent hot air recirculation between the front and rear of the rack. Critical for maintaining hot/cold aisle separation.

Related:Hot AisleCold Aisle
C
9 terms

Cage

Data Center

A physically secured, fenced area within a colocation facility allocated to a single tenant. Provides dedicated space with controlled access.

Related:SuiteCabinetColocation

CapEx (Capital Expenditure)

Cloud

Upfront investment in physical assets (servers, network equipment, facility). On-premises infrastructure is primarily CapEx.

Related:OpExTCODepreciation

CDU (Coolant Distribution Unit)

Cooling

A device that conditions and distributes cooling liquid to server cold plates in a direct liquid cooling system. Manages temperature, pressure, and flow rate.

Related:DLCManifoldChilled Water

CISSP (Certified Information Systems Security Professional)

Cybersecurity

The gold standard cybersecurity certification issued by (ISC)². Requires 5 years of experience and covers 8 security domains.

(ISC)²
Related:CISMSecurity Architecture

Colocation (Colo)

Data Center

A data center facility where multiple customers house their own servers and networking equipment in a shared facility. The facility provides power, cooling, physical security, and network connectivity.

Related:CageSuiteCross-Connect

Concurrently Maintainable

Data Center

A Tier III data center characteristic where any component can be maintained or replaced without shutting down the IT load. Requires redundant paths for power and cooling.

Uptime Institute Tier Standard
Related:Tier IIIFault Tolerant

CRAC (Computer Room Air Conditioner)

Cooling

A self-contained precision cooling unit with its own refrigeration circuit. Cools air by passing it over a direct expansion (DX) coil. Common in smaller data centers and edge deployments.

Related:CRAHAHUPrecision Cooling

CRAH (Computer Room Air Handler)

Cooling

A precision cooling unit that uses chilled water (from a central chiller plant) rather than a self-contained refrigeration circuit. More efficient at scale than CRAC units.

Related:CRACChillerChilled Water

CUDA (Compute Unified Device Architecture)

AI & GPU

NVIDIA's parallel computing platform and programming model that enables GPU acceleration for general-purpose computing. The foundation of the NVIDIA AI software ecosystem.

Related:Tensor CorecuDNNNCCL
D
5 terms

DCIM (Data Center Infrastructure Management)

Data Center

Software platform that monitors, manages, and optimizes data center infrastructure including power, cooling, space, and assets in real time.

Related:BMSIPAMCMDB

DGX (Data Center GPU)

AI & GPU

NVIDIA's purpose-built AI infrastructure systems. DGX H100 contains 8x H100 SXM5 GPUs with NVLink interconnect, 640GB HBM3 memory, and dual 400GbE/InfiniBand connectivity.

Related:HGXNVLinkInfiniBand

Direct Liquid Cooling (DLC)

Cooling

A cooling method where cold plates are attached directly to heat-generating components (CPUs, GPUs) and liquid is circulated through them. Supports 100+ kW/rack. Required for NVIDIA H100 SXM5 and GB200.

Related:Immersion CoolingCDUPower Density

Double-Conversion UPS

Power Systems

A UPS topology where incoming AC power is converted to DC, then back to AC. Provides complete electrical isolation and the highest power quality. The standard for mission-critical data centers.

IEC 62040-3
Related:UPSLine-InteractiveBypass

DWDM (Dense Wavelength Division Multiplexing)

Networking

A fiber optic technology that multiplexes multiple optical signals onto a single fiber using different wavelengths. Enables 100+ Gbps over long distances.

ITU-T G.694.1
Related:Dark FiberOptical Transport
E
2 terms

ECMP (Equal-Cost Multi-Path)

Networking

A routing strategy that distributes traffic across multiple equal-cost paths simultaneously, increasing bandwidth and providing redundancy. Essential in spine-leaf architectures.

RFC 2992
Related:Spine-LeafBGPLoad Balancing

Economizer (Free Cooling)

Cooling

A cooling mode where outdoor air or water is used directly for cooling without mechanical refrigeration, reducing energy consumption. Air-side economizers use outdoor air directly; water-side use cooling towers.

ASHRAE TC 9.9
Related:PUEChillerASHRAE
F
3 terms

Fault Tolerant

Data Center

A Tier IV data center characteristic where any single failure — including a complete path failure — does not interrupt IT operations. Requires 2N+1 minimum redundancy.

Uptime Institute Tier Standard
Related:Tier IV2N RedundancyConcurrently Maintainable

FISMA (Federal Information Security Management Act)

Cybersecurity

U.S. federal law requiring federal agencies and contractors to implement information security programs. Compliance is required for all federal IT systems.

44 U.S.C. § 3551
Related:FedRAMPNIST SP 800-53ATO

Floor Loading

Data Center

The maximum weight per square foot (or kg/m²) a raised floor or concrete slab can support. Critical for high-density AI racks. Standard raised floor: 1,000–2,000 lbs/sq ft. AI racks may require 3,000+ lbs/sq ft.

Related:Power DensityRaised Floor
G
2 terms

Generator (Standby)

Power Systems

A diesel or natural gas engine-driven alternator that provides backup power during utility outages. Typically starts within 10 seconds and reaches full load within 30 seconds.

NFPA 110
Related:ATSUPSFuel Storage

GPU (Graphics Processing Unit)

AI & GPU

A massively parallel processor originally designed for graphics rendering, now the primary compute engine for AI/ML workloads. Contains thousands of cores optimized for matrix operations.

Related:Tensor CoreCUDAHBM
H
5 terms

HBM (High Bandwidth Memory)

AI & GPU

A high-speed, high-capacity memory technology stacked directly on the GPU die. HBM3e in H200 provides 4.8 TB/s memory bandwidth, critical for large model inference.

JEDEC HBM Standard
Related:VRAMMemory BandwidthGPU

Hot Aisle / Cold Aisle

Data Center

A data center layout where server racks are arranged so cold air intakes face a cold aisle and hot air exhausts face a hot aisle. Prevents hot and cold air mixing, improving cooling efficiency.

Related:ContainmentCRACPUE

Hot Aisle Containment (HAC)

Cooling

A physical enclosure around the hot aisle that captures server exhaust air and directs it back to cooling units, preventing mixing with cold supply air.

Related:Cold Aisle ContainmentPUECRAC

Hybrid Cloud

Cloud

An IT architecture combining on-premises infrastructure with public cloud services, connected by a private network or VPN. Enables workload portability and burst capacity.

Related:Multi-CloudPrivate CloudEdge

Hyperscale

Data Center

Data centers exceeding 100MW of IT load, typically operated by cloud providers (AWS, Azure, GCP, Meta, Google). Characterized by extreme standardization, automation, and economies of scale.

Related:ColocationEdge
I
3 terms

Immersion Cooling

Cooling

A cooling method where IT equipment is submerged in a thermally conductive but electrically non-conductive liquid. Single-phase uses mineral oil; two-phase uses fluorocarbon fluids that boil and condense. Supports 200+ kW/rack.

Related:DLCPUEPower Density

InfiniBand

AI & GPU

A high-performance networking technology providing ultra-low latency (600ns) and high bandwidth (NDR: 400Gb/s per port). The dominant interconnect for large GPU training clusters.

InfiniBand Trade Association
Related:RDMARoCEv2NDR

ISO 27001

Cybersecurity

The international standard for information security management systems (ISMS). Provides a framework for establishing, implementing, and maintaining information security.

ISO/IEC 27001:2022
Related:SOC 2NISTISMS
K
1 term

kVA (Kilovolt-Ampere)

Power Systems

Unit of apparent power in AC circuits. Actual power (kW) = kVA × Power Factor. UPS systems are rated in kVA. A 100 kVA UPS at 0.9 PF delivers 90 kW.

Related:kWPower FactorUPS
L
1 term

LLM (Large Language Model)

AI & GPU

A deep learning model trained on massive text datasets with billions to trillions of parameters. Examples: GPT-4, LLaMA, Claude, Gemini. Infrastructure requirements scale with parameter count.

Related:TransformerTrainingInference
M
3 terms

MTBF (Mean Time Between Failures)

Data Center

Statistical measure of the average time between system failures. Used in reliability engineering to predict component and system failure rates.

Related:MTTRAvailabilityRedundancy

MTTR (Mean Time to Repair)

Data Center

Average time required to restore a failed component or system to operational status. Directly impacts overall system availability.

Related:MTBFRTORPO

Multi-Cloud

Cloud

A strategy using services from multiple cloud providers (AWS + Azure + GCP) to avoid vendor lock-in, optimize cost, or meet data sovereignty requirements.

Related:Hybrid CloudCloud Broker
N
8 terms

N+1 Redundancy

Power Systems

A redundancy configuration where N components are required for operation and one additional component is available as a spare. If any single component fails, the spare takes over.

Uptime Institute Tier Standard
Related:2N RedundancyN+2Redundancy

NCCL (NVIDIA Collective Communications Library)

AI & GPU

NVIDIA's library for multi-GPU and multi-node collective communication operations (AllReduce, AllGather, etc.) used in distributed training.

Related:InfiniBandRoCEv2Distributed Training

NDR (Next Data Rate)

Networking

InfiniBand's current generation standard providing 400 Gb/s per port (800 Gb/s bidirectional). Successor to HDR (200 Gb/s). The standard for large AI training clusters.

InfiniBand Trade Association
Related:InfiniBandHDRRDMA

NETA (InterNational Electrical Testing Association)

Standards

The organization that establishes standards for electrical testing and maintenance. NETA MTS (Maintenance Testing Specifications) is the standard for electrical acceptance testing in data centers.

ANSI/NETA MTS
Related:CommissioningAcceptance TestingOSHA

NIST SP 800-53

Cybersecurity

NIST's catalog of security and privacy controls for federal information systems. The foundation of U.S. government cybersecurity compliance.

NIST SP 800-53 Rev. 5
Related:FISMAFedRAMPZero-Trust Architecture

NVLink

AI & GPU

NVIDIA's high-speed GPU-to-GPU interconnect within a single node. NVLink 4.0 provides 900 GB/s bidirectional bandwidth between GPUs, far exceeding PCIe 5.0's 128 GB/s.

Related:NVSwitchDGXGPU

NVMe (Non-Volatile Memory Express)

Storage

A storage protocol designed specifically for flash storage, providing dramatically lower latency (microseconds vs milliseconds for SATA/SAS) and higher IOPS.

NVM Express Base Specification
Related:PCIeAll-Flash ArrayU.2

NVSwitch

AI & GPU

NVIDIA's switching chip that enables all-to-all NVLink connectivity between all GPUs in a DGX system. NVSwitch 3.0 in DGX H100 provides 900 GB/s per GPU.

Related:NVLinkDGX
O
2 terms

Object Storage

Storage

A storage architecture that manages data as objects (vs files or blocks). Highly scalable, ideal for unstructured data (AI datasets, backups, archives). Examples: AWS S3, MinIO, Ceph.

Related:S3Block StorageFile Storage

OpEx (Operational Expenditure)

Cloud

Ongoing costs for running a business (cloud subscriptions, maintenance contracts, staffing). Cloud infrastructure is primarily OpEx.

Related:CapExTCOSaaS
P
3 terms

Parallel File System

Storage

A distributed file system that stores data across multiple storage nodes simultaneously, providing aggregate bandwidth that scales with the number of nodes. Examples: GPFS (IBM Spectrum Scale), Lustre, WEKA. Required for large AI training clusters.

Related:NFSPOSIXAI Storage

PDU (Power Distribution Unit)

Power Systems

A device that distributes electrical power to multiple IT equipment outlets within a rack or row. Types range from basic (passive) to intelligent (monitored, switched, with outlet-level metering).

Related:Rack PDUBuswayATS

PUE (Power Usage Effectiveness)

Data Center

The ratio of total data center power consumption to IT equipment power consumption. PUE = Total Facility Power ÷ IT Equipment Power. A PUE of 1.0 is theoretical perfection; world-class facilities achieve 1.05–1.12.

The Green Grid / ISO/IEC 30134-2
Related:DCiEWUECUE
R
3 terms

Raised Floor

Data Center

A modular floor system elevated above the structural slab, creating a plenum for cable management and underfloor air distribution. Standard height: 12–24 inches.

Related:Hot AisleCold AisleBlanking Panel

RDMA (Remote Direct Memory Access)

AI & GPU

A technology allowing direct memory access from one computer's memory to another's without involving the CPU. Critical for low-latency GPU cluster communication.

Related:InfiniBandRoCEv2NCCL

RoCEv2 (RDMA over Converged Ethernet v2)

AI & GPU

/ROH-see-vee-two/

RDMA protocol running over standard Ethernet. Requires lossless network (PFC + ECN). Lower cost than InfiniBand but higher latency.

Related:RDMAInfiniBandPFC
S
4 terms

SDN (Software-Defined Networking)

Networking

A network architecture approach that separates the control plane from the data plane, enabling centralized, programmable network management.

ONF TR-521
Related:NFVOpenFlowSpine-Leaf

SOC 2 Type II

Cybersecurity

An auditing standard developed by the AICPA that evaluates a service organization's controls related to security, availability, processing integrity, confidentiality, and privacy over a period of time (typically 6–12 months).

AICPA TSC
Related:ISO 27001FISMATrust Services Criteria

Spine-Leaf

Networking

A two-tier network topology where every leaf switch connects to every spine switch, providing predictable latency and easy horizontal scaling. The standard architecture for modern data centers.

Related:ECMPOversubscriptionClos Network

Static Transfer Switch (STS)

Power Systems

An electronic switching device that transfers load between two power sources in less than 4 milliseconds — fast enough to prevent IT equipment from detecting the transfer.

IEC 62310
Related:ATSUPSRedundancy
T
4 terms

TCO (Total Cost of Ownership)

Cloud

The complete cost of an asset over its useful life, including acquisition, operation, maintenance, and disposal. Essential for build vs. buy vs. cloud decisions.

Related:CapExOpExROI

Tensor Core

AI & GPU

Specialized processing units within NVIDIA GPUs designed specifically for matrix multiply-accumulate operations (the core computation in neural networks). H100 has 528 Tensor Cores.

Related:CUDAGPUFP8

TIA-942

Standards

A telecommunications infrastructure standard for data centers published by the Telecommunications Industry Association. Defines Rated-1 through Rated-4 classifications (similar to but distinct from Uptime Institute Tiers).

TIA-942-B
Related:Uptime InstituteTier Classification

Tier I / II / III / IV

Data Center

The Uptime Institute's data center classification system based on redundancy, availability, and fault tolerance. Tier I = 99.671% availability; Tier IV = 99.995% availability.

Uptime Institute Tier Standard
Related:Uptime InstituteConcurrently MaintainableFault Tolerant
U
2 terms

UPS (Uninterruptible Power Supply)

Power Systems

A device that provides emergency power to IT equipment when the main power source fails. Bridges the gap between utility failure and generator startup (typically 10–30 seconds).

IEC 62040
Related:BatteryGeneratorATS

Uptime Institute

Standards

The global authority on data center performance and reliability. Publishes the Tier Standard and provides independent certification of Tier I–IV data centers.

Related:Tier I / II / III / IVConcurrently MaintainableFault Tolerant
V
2 terms

VRAM (Video RAM)

AI & GPU

The dedicated memory on a GPU. Determines maximum model size for inference (model must fit in VRAM). H100: 80GB HBM3; H200: 141GB HBM3e; B200: 192GB HBM3e.

Related:HBMGPULLM

VXLAN (Virtual Extensible LAN)

Networking

A network virtualization technology that encapsulates Layer 2 frames within UDP packets, enabling Layer 2 networks to span Layer 3 boundaries. Used for multi-tenant data center networks.

RFC 7348
Related:SDNOverlay NetworkEVPN
Z
1 term

Zero-Trust Architecture

Cybersecurity

A security model based on the principle "never trust, always verify." Eliminates implicit trust based on network location; every access request is authenticated, authorized, and continuously validated.

NIST SP 800-207
Related:NIST SP 800-207MicrosegmentationIdentity-Aware Proxy

Missing a term?

Our engineering team reviews and adds new definitions monthly.

Submit a Term