Chip registry
Reference-design clusters
The layer between one chip and one campus. Buyers order rack-scale reference designs (NVL72, HGX, OAM platforms, TPU pods, Trainium UltraServers). Every cluster below links back to the chips it contains and the campuses it lands in.
NVIDIA GB200 NVL72
rack72 Blackwell GPUs + 36 Grace CPUs in one liquid-cooled rack, sharing a single NVLink 5 domain. The current flagship training rack.
36 chips per instance · 132 kW · water-cooled · GA 2024-12-01
NVIDIA HGX H100 (8-GPU baseboard)
platformThe 8-GPU H100 SXM5 baseboard that anchored the entire 2023-2024 AI-training build-out. The generation the world learned to buy in bulk.
8 chips per instance · 10.2 kW · air-cooled · GA 2022-10-13
NVIDIA HGX H200 (8-GPU baseboard)
platformH100's memory-refresh sibling: same footprint, 141 GB HBM3e per GPU (+76%) and 4.8 TB/s bandwidth. Inference-oriented shops upgraded here.
8 chips per instance · 10.2 kW · air-cooled · GA 2024-03-01
NVIDIA HGX B200 (8-GPU baseboard)
platformThe Blackwell-generation 8-GPU baseboard for buyers who wanted a familiar HGX form-factor instead of committing to NVL72's rack-scale.
8 chips per instance · 14.3 kW · hybrid-cooled · GA 2024-11-01
NVIDIA DGX SuperPOD H100 (32-node reference)
supercluster32 DGX H100 nodes = 256 GPUs, 32 InfiniBand switches, one full-fat rail-optimised training pod. The AI startup unit-of-currency in 2023.
256 chips per instance · 480 kW · air-cooled · GA 2023-01-01
AMD Instinct MI300X Platform (8-GPU OAM baseboard)
platformAMD's 8-way OAM answer to HGX H100/H200: 1.5 TB HBM3 total, priced to unseat NVIDIA on memory-bound inference. The Meta / Microsoft / Oracle purchase.
8 chips per instance · 6 kW · air-cooled · GA 2024-01-01
AMD Instinct MI325X Platform (8-GPU OAM baseboard)
platformThe MI325X refresh: 256 GB HBM3e per GPU, 2 TB per baseboard. AMD's tightest inference-per-dollar pitch heading into 2025.
8 chips per instance · 8 kW · air-cooled · GA 2025-01-01
Google TPU v5p Pod (8,960 chips)
superclusterGoogle's single-pod ceiling under v5p: 8,960 accelerators in one all-to-all ICI domain. The unit-of-scale for Gemini-generation training.
8,960 chips per instance · water-cooled · GA 2024-01-01
Google TPU v6e (Trillium) Pod-256
podThe inference-optimised Trillium pod slice at 256 chips. Google Cloud's default Gemini-serving unit.
256 chips per instance · water-cooled · GA 2024-12-01
AWS Trainium 2 UltraServer
ultraserverAWS's answer to NVL72: 64 Trainium 2 chips lashed together with NeuronLink. Anthropic's Project Rainier compute unit.
64 chips per instance · water-cooled · GA 2025-01-01
Huawei Atlas 900 A3 SuperCluster (Ascend 910B)
superclusterThe Chinese ceiling under US export controls: 8,192 domestically fabbed Ascend 910B accelerators in one all-to-all HCCS fabric.
8,192 chips per instance · water-cooled · GA 2025-01-01
See also: every chip on record · every foundry that fabs them · head-to-head chip comparisons.