GB300-NVL72Blackwell Ultra (Grace Blackwell)
NVIDIA GB300 NVL72
72 Blackwell Ultra GPUs and 36 Grace CPUs in one liquid-cooled rack and one NVLink domain, on reserved terms.
In one NVLink domain
72 GPUs
NVIDIA reference figure
- 1. Power shelf · 8 per rack
- 2. Management switch · 2 per rack
- 3. Compute tray · 18 × 1RU, 2 Grace CPUs and 4 GPUs each
- 4. NVLink switch tray · 9 × 1RU
- 5. NVLink switch chip · 2 per switch tray
- 6. Grace CPU · 2 per compute tray, 36 per rack
- 7. Blackwell Ultra GPU · 4 per compute tray, 72 per rack
- 8. ConnectX-8 SuperNIC, 800 Gb/s · 4 per compute tray
- 9. Liquid-cooling manifold · rear, supply and return
- 10. Bus bar · rear, fed by the power shelves
Point at a part or a row. The other one lights.
GB300 NVL72 specifications.
| Parameter | Value | Scope | Source |
|---|---|---|---|
| Memory | |||
| Memory | 279 GB HBM3E | Per GPU | [1] |
| Memory bandwidth | 8 TB/s | Per GPU | [1] |
| GPU memory in the rack | 20 TB HBM3E | Per rack | [3] |
| GPU memory bandwidth | Up to 576 TB/s | Per rack | [3] |
| CPU memory | 17 TB LPDDR5X · 14 TB/s | Per rack | [3] |
| Fast memory | 37 TB | Per rack | [3] |
| Compute, peak · dense / sparse | |||
| FP4 Tensor Core | 15 / 20 PFLOPS | Per GPU | [1] |
| FP8 Tensor Core | 5 / 10 PFLOPS | Per GPU | [1] |
| BF16 Tensor Core | Not published | Per GPU | |
| FP4 Tensor Core | 1,080 / 1,440 PFLOPS | Per rack | [3] |
| FP8 Tensor Core | 360 / 720 PFLOPS | Per rack | [3] |
| Interconnect | |||
| NVLink 5 | 1.8 TB/s | Per GPU | [2] |
| NVLink domain | 72 GPUs in one NVLink domain · 130 TB/s across the rack | Per rack | [3] |
| Host interface | PCIe Gen6 · 256 GB/s | Per GPU | [1] |
| Scale-out network | 4 × ConnectX-8 800G per compute tray · 800 Gb/s per GPU | Per compute tray | [5] |
| Rack | |||
| GPUs | 72 | Per rack | [4] |
| Grace CPUs | 36 | Per rack | [4] |
| Compute trays | 18 · 2 Grace CPUs and 4 GPUs each | Per rack | [4] |
| NVLink switch trays | 9 | Per rack | [4] |
| Power and form factor | |||
| Power | Up to 1,400 W, configurable | Per GPU | [1] |
| Form factor | Liquid-cooled rack; Superchip modules (1 Grace CPU, 2 GPUs) | [1] | |
| Cooling | Fully liquid-cooled rack | Per rack | [4] |
Specifications are NVIDIA reference figures.
Best for
- Reasoning and test-time-scaling inference
- Large-scale training
- Video generation
Look elsewhere for
- FP64 HPC
Notes on the figures
- Memory: 279 GB per GPU (datasheet). The launch figure of up to 288 GB is not used.
- Reserved capacity only. Not offered on demand or as spot, and not a Workspaces or Sandboxes shape.
Reserved capacity, quoted per rack.
Rack-scale systems are reserved capacity. Talk to us about term, delivery and configuration.
-
Reserved
Contact sales
Take-or-pay under a signed order form. No early cancellation without penalty (Terms §3).
-
Term and delivery
Set in the order form
Term, delivery and configuration are agreed for each rack before the order is signed.
-
On-demand and spot
Not offered
Reserved capacity never runs as spot.
GB300 NVL72 listing
- Status
- By request
- Runs as
- Reserved racks
- Capacity
- Reserved
- Tied to
- One GPU type, one region, one InfiniBand fabric
One rack, one NVLink domain.
18 compute trays, each with 2 Grace CPUs and 4 GPUs, and 9 NVLink switch trays join 72 GPUs in one NVLink domain.
One GB300 NVL72 rack holds 18 compute trays. Each compute tray holds 4 GPUs and 2 Grace CPUs, which makes 72 GPUs and 36 Grace CPUs. 9 NVLink switch trays join all 72 GPUs in one NVLink domain, 130 TB/s across the rack. The figure shows counts, not positions in the rack. Connections: Compute tray to NVLink switch tray: NVLink 5; NVLink switch tray to One NVLink domain: 72 GPUs; One NVLink domain to In the rack.
- GPUs
- 4
- Grace CPUs
- 2
18 trays4 × ConnectX-8 800G per compute tray800 Gb/s per GPU
- per GPU
- 1.8 TB/s
9 trays- GPUs
- 72
- Grace CPUs
- 36
- NVLink 5
- 130 TB/s across the rack
- GPU memory
- 20 TB HBM3E
- GPU bandwidth
- Up to 576 TB/s
- CPU memory
- 17 TB LPDDR5X · 14 TB/s
- Fast memory
- 37 TB
- FP4
- 1,080 / 1,440 PFLOPS
- FP8
- 360 / 720 PFLOPS
Tensor Core peak, dense / sparse.
- cooling
- Fully liquid-cooled rack
- GPUs
- 4
- Grace CPUs
- 2
18 trays4 × ConnectX-8 800G per compute tray800 Gb/s per GPU
- per GPU
- 1.8 TB/s
9 trays- GPUs
- 72
- Grace CPUs
- 36
- NVLink 5
- 130 TB/s across the rack
- GPU memory
- 20 TB HBM3E
- GPU bandwidth
- Up to 576 TB/s
- CPU memory
- 17 TB LPDDR5X · 14 TB/s
- Fast memory
- 37 TB
- FP4
- 1,080 / 1,440 PFLOPS
- FP8
- 360 / 720 PFLOPS
- cooling
- Fully liquid-cooled rack
- GPUs
- 4
- Grace CPUs
- 2
18 trays4 × ConnectX-8 800G per compute tray800 Gb/s per GPU
- per GPU
- 1.8 TB/s
9 trays- GPUs
- 72
- Grace CPUs
- 36
- NVLink 5
- 130 TB/s across the rack
- GPU memory
- 20 TB HBM3E
- GPU bandwidth
- Up to 576 TB/s
- CPU memory
- 17 TB LPDDR5X · 14 TB/s
- Fast memory
- 37 TB
- FP4
- 1,080 / 1,440 PFLOPS
- FP8
- 360 / 720 PFLOPS
- cooling
- Fully liquid-cooled rack
NVIDIA GB300 NVL72 on Fantasti: fantasti.ai/gpus/gb300-nvl72 · Request capacity: fantasti.ai/contact
- Sheet
- 01 / 01
- Title
- GB300 NVL72 datasheet
- Figures
- NVIDIA reference
- Reviewed
- 2026-10-10
Request GB300 NVL72 capacity.
Rack-scale systems are reserved capacity. Talk to us about term, delivery and configuration.