HomeSystems › NVIDIA GB200 NVL72
rack-scale · Blackwell

NVIDIA GB200 NVL72

NVIDIA GB200 NVL72 rack: 72 Blackwell GPUs and 36 Grace CPUs as one NVLink domain, 13.5TB HBM3e, liquid-cooled, around 120kW per rack. What it is for, what it needs, and how to get quotes.

72× B20036 Grace CPUs
13.5 TBHBM3e in one NVLink domain
130 TB/sNVLink bandwidth
~120 kWper rack, liquid-cooled

One rack as one GPU

GB200 NVL72 connects 72 Blackwell GPUs and 36 Grace CPUs into a single NVLink domain, so a trillion-parameter model sees the whole rack as one accelerator. NVIDIA's own figure is roughly 30 times the real-time inference throughput of an equivalent H100 cluster on the largest models, at a fraction of the energy per token. It ships only as a liquid-cooled rack, built by Supermicro and other partners to NVIDIA's reference design.

Who it is for

  • Frontier-scale training and inference on models above the 400B class.
  • Providers selling Blackwell capacity by the hour.
  • Nobody running 70B models: an H200 node does that for far less.

The facility question comes first

  • Around 120kW per rack means most existing European colocation halls cannot host it without an upgrade. We check this before quoting.
  • Facility water or CDU capacity is mandatory.
  • Allocation is still tight; lead times are quoted case by case.

Cloud pricing for B200 by the hour is on our price table if you want the capacity without the building work.

Get three quotes for the GB200 NVL72

Comparable pricing and lead times from authorised integrators, plus a cloud and hosted alternative so you can see the trade-off.

Request quotes