B200Blackwell
NVIDIA B200
Large-model training and inference on NVIDIA Blackwell, with NVLink 5 at 1.8 TB/s per GPU.
HBM3E per GPU
180 GB
NVIDIA reference figure
- 1. SXM6 module · 8 per baseboard, four per side
- 2. Heatsink · one per module; this one is drawn lifted
- 3. B200 GPU · one per module, under the heatsink
- 4. NVLink Switch chip · 2 per baseboard, in the board centre
- 5. Baseboard · joins the 8 GPUs in one NVLink domain
Point at a part or a row. The other one lights.
B200 specifications.
| Parameter | Value | Scope | Source |
|---|---|---|---|
| Memory | |||
| Memory | 180 GB HBM3E | Per GPU | [2] |
| Memory bandwidth | 7.7 TB/s | Per GPU | [2] |
| Memory on one board | 1,440 GB · 8 × 180 GB | Per HGX board | [2] |
| Compute, peak · dense / sparse | |||
| FP4 Tensor Core | 9 / 18 PFLOPS | Per GPU | [2] |
| FP8 Tensor Core | 4.5 / 9 PFLOPS | Per GPU | [2] |
| BF16 Tensor Core | 2.25 / 4.5 PFLOPS | Per GPU | [2] |
| Interconnect | |||
| NVLink 5 | 1.8 TB/s | Per GPU | [3] |
| NVLink domain | 8 GPUs per HGX B200 board · 14.4 TB/s aggregate | Per HGX board | [5] |
| Host interface | PCIe Gen5 · 128 GB/s | Per GPU | [2] |
| Power and form factor | |||
| Power | Up to 1,000 W, configurable | Per GPU | [2] |
| Form factor | SXM module (SXM6) on the HGX B200 8-GPU baseboard | [2] | |
| Board | 8 SXM6 GPUs per HGX B200 board · 2 NVLink Switch chips | Per HGX board | [5] |
Specifications are NVIDIA reference figures.
Best for
- Large-model training
- High-throughput inference
- HPC
Notes on the figures
- Memory: 180 GB per GPU (NVIDIA datasheet; DGX B200 lists 1,440 GB per 8 GPUs). The older 192 GB figure is not used.
- Bandwidth: 7.7 TB/s per GPU (datasheet). The DGX B200 64 TB/s system figure is not used.
- BF16: 18 / 36 PFLOPS dense / sparse per 8-GPU board, shown per GPU.
On-demand, spot or reserved.
Metered by the second, billed hourly. Reserved rates are quoted per term.
-
On-demand
$9.35 /GPU·hr
- Per GPU-second
- $0.002597
- Per 8 GPUs, per hour
- $74.80
- Per GPU-month, 730 h
- $6,825.50
Preview list price. It applies when your account opens.
-
Spot
From $1.09 /GPU·hr
Work that checkpoints and can wait. Follow the current spot price, or set a maximum hourly price. If the spot price rises above your max, the workload may stop.
"From" is the lowest spot price for that GPU. The spot price moves with supply and demand. The spot price always stays below the on-demand rate for the same GPU.
Spot price moves with supply and demand. Follow it, or set a max price per GPU-hour: you pay the spot price, and nodes stop if it rises above your max. A max price limits what you pay. It does not reserve capacity. To guarantee capacity, reserve it. How spot works
-
Reserved
Request a quote
Capacity you need on a date, guaranteed. 1 month, 6 months or 12+ months.
Ways to pay Prepaid credit, pay-as-you-go and Enterprise. Each one covers on-demand, spot and reserved capacity. How billing works
| Term | Commitment | Capacity | Rate |
|---|---|---|---|
| On-demandNone · Not held | None | Not held | $10.45 /GPU·hr |
| 1 monthTake-or-pay · Held for the term, on one fabric | Take-or-pay | Held for the term, on one fabric | Request a quote |
| 6 monthsTake-or-pay · Held for the term, on one fabric | Take-or-pay | Held for the term, on one fabric | Request a quote |
| 12+ monthsTake-or-pay · Held for the term, on one fabric | Take-or-pay | Held for the term, on one fabric | Request a quote |
| Term | Commitment | Capacity | Rate |
|---|---|---|---|
| On-demandNone · Not held | None | Not held | $9.35 /GPU·hr |
| 1 monthTake-or-pay · Held for the term, on one fabric | Take-or-pay | Held for the term, on one fabric | Request a quote |
| 6 monthsTake-or-pay · Held for the term, on one fabric | Take-or-pay | Held for the term, on one fabric | Request a quote |
| 12+ monthsTake-or-pay · Held for the term, on one fabric | Take-or-pay | Held for the term, on one fabric | Request a quote |
| Term | Commitment | Capacity | Rate |
|---|---|---|---|
| On-demandNone · Not held | None | Not held | $5.94 /GPU·hr |
| 1 monthTake-or-pay · Held for the term, on one fabric | Take-or-pay | Held for the term, on one fabric | Request a quote |
| 6 monthsTake-or-pay · Held for the term, on one fabric | Take-or-pay | Held for the term, on one fabric | Request a quote |
| 12+ monthsTake-or-pay · Held for the term, on one fabric | Take-or-pay | Held for the term, on one fabric | Request a quote |
| Term | Commitment | Capacity | Rate |
|---|---|---|---|
| On-demandNone · Not held | None | Not held | $5.40 /GPU·hr |
| 1 monthTake-or-pay · Held for the term, on one fabric | Take-or-pay | Held for the term, on one fabric | Request a quote |
| 6 monthsTake-or-pay · Held for the term, on one fabric | Take-or-pay | Held for the term, on one fabric | Request a quote |
| 12+ monthsTake-or-pay · Held for the term, on one fabric | Take-or-pay | Held for the term, on one fabric | Request a quote |
| Term | Commitment | Capacity | Rate |
|---|---|---|---|
| On-demandNone · Not held | None | Not held | $2.16 /GPU·hr |
| 1 monthTake-or-pay · Held for the term | Take-or-pay | Held for the term | Request a quote |
| 6 monthsTake-or-pay · Held for the term | Take-or-pay | Held for the term | Request a quote |
| 12+ monthsTake-or-pay · Held for the term | Take-or-pay | Held for the term | Request a quote |
| Term | Commitment | Capacity | Rate |
|---|---|---|---|
| On-demandNone · Not held | None | Not held | $1.86 /GPU·hr |
| 1 monthTake-or-pay · Held for the term | Take-or-pay | Held for the term | Request a quote |
| 6 monthsTake-or-pay · Held for the term | Take-or-pay | Held for the term | Request a quote |
| 12+ monthsTake-or-pay · Held for the term | Take-or-pay | Held for the term | Request a quote |
List price for the base shape: 1 GPU, 8 vCPU, 32 GiB. Larger shapes are quoted.
Reserved rates are quoted per term. Reservations are tied to a GPU type, a region and one InfiniBand fabric. Take-or-pay under a signed order form. No early cancellation without penalty (Terms §3).
Runs as instances and as cluster nodes.
HGX B200: 8 GPUs, 2 NVLink Switch chips, 1,440 GB of HBM3E per node.
B200 runs as a 1-GPU instance (20 vCPU · 224 GiB) or as an 8-GPU node (160 vCPU · 1,792 GiB). 8-GPU nodes join one InfiniBand fabric at 400 Gb/s per GPU, 3.2 Tb/s per node. Host network adapter: 400 Gb/s per machine. Connections: 8 GPUs to One InfiniBand fabric, 8 links: 8 × 400 Gb/s; More 8-GPU nodes to One InfiniBand fabric; 1 GPU to Host network adapter.
- vCPU
- 20
- RAM
- 224 GiB
1 GPU- rating
- 400 Gb/s
- scope
- per machine
- vCPU
- 160
- RAM
- 1,792 GiB
- CPU
- Intel Xeon Platinum
- 8580 or 8570
8 GPUs · NVLink 5- per GPU
- 400 Gb/s
- per node
- 3.2 Tb/s
Every node of one cluster sits on the same fabric. A cluster is placed onone fabric and stays there.
- vCPU
- 20
- RAM
- 224 GiB
1 GPU- rating
- 400 Gb/s
- scope
- per machine
- vCPU
- 160
- RAM
- 1,792 GiB
- CPU
- Intel Xeon
- Platinum 8580
- or 8570
8 GPUs · NVLink 5- per GPU
- 400 Gb/s
- per node
- 3.2 Tb/s
- vCPU
- 20
- RAM
- 224 GiB
1 GPU- rating
- 400 Gb/s
- scope
- per machine
- vCPU
- 160
- RAM
- 1,792 GiB
- CPU
- Intel Xeon Platinum 8580 or 8570
8 GPUs · NVLink 5- per GPU
- 400 Gb/s
- per node
- 3.2 Tb/s
B200 listing.
CPU, RAM, local NVMe and the fabric vary by pool and are reported at placement. GPUs and InfiniBand adapters are passed through whole to one machine.
- Status
- Private Preview
- Runs as
- Single instances and GPU Clusters
- Capacity
- On-demand · Spot · Reserved
- Metering
- Metered by the second, billed hourly
| Parameter | Value | Scope |
|---|---|---|
| 1 GPU | 20 vCPU · 224 GiB | Per machine |
| 8 GPUs | 160 vCPU · 1,792 GiB | Per machine |
| Host CPU | Intel Xeon Platinum 8580 or 8570 | Per 8-GPU node |
| InfiniBand | 400 Gb/s | Per GPU |
| InfiniBand | 3.2 Tb/s | Per 8-GPU node |
| Host network adapter | 400 Gb/s | Per machine |
NVIDIA B200 on Fantasti: fantasti.ai/gpus/b200 · Request capacity: fantasti.ai/contact
- Sheet
- 01 / 01
- Title
- B200 datasheet
- Figures
- NVIDIA reference
- Reviewed
- 2026-10-10
Request B200 capacity.
Tell us the count, the term and where it should run. We reply to a reviewed request, usually within one business day. A reserved term is quoted.