RTX-PRO-6000Blackwell
NVIDIA RTX PRO 6000 Blackwell Server Edition
Inference, agentic AI, simulation and rendering on PCIe servers, with 96 GB of GDDR7 per card.
GDDR7 with ECC per GPU
96 GB
NVIDIA reference figure
- 1. PCIe x16 edge connector · PCIe 5.0
- 2. Passive heatsink · no fan on the card; the chassis moves the air
- 3. Cover · drawn cut away to show the fins
- 4. GPU, 96 GB GDDR7 with ECC · under the heatsink
- 5. Bracket · dual-slot, full height
Point at a part or a row. The other one lights.
RTX PRO 6000 specifications.
| Parameter | Value | Scope | Source |
|---|---|---|---|
| Memory | |||
| Memory | 96 GB GDDR7 with ECC | Per GPU | [1] |
| Memory bandwidth | 1.597 TB/s | Per GPU | [1] |
| Compute, peak · dense / sparse | |||
| FP4 Tensor Core | 4 PFLOPS (peak, sparse) | Per GPU | [2] |
| FP8 Tensor Core | 2 PFLOPS (peak, sparse) | Per GPU | [2] |
| BF16 Tensor Core | Not published | Per GPU | |
| Interconnect | |||
| NVLink | None | ||
| Host interface | PCIe 5.0 x16 | Per GPU | [1] |
| Power and form factor | |||
| Power | Up to 600 W, configurable | Per GPU | [1] |
| Form factor | PCIe card: dual-slot FHFL air-cooled, or single-slot FHXL liquid-cooled | [4] | |
Specifications are NVIDIA reference figures.
Best for
- Inference and agentic AI
- Physical AI and simulation
- Rendering and video (4 NVENC, 4 NVDEC)
Look elsewhere for
- Multi-node training that needs InfiniBand
Notes on the figures
- Compute: peak, sparse. Dense FP8 and FP4 are not published by NVIDIA and are left blank.
- No NVLink and no InfiniBand: cannot join a multi-node GPU cluster. MIG: up to 4 instances of 24 GB.
On-demand, spot or reserved.
Metered by the second, billed hourly. Reserved rates are quoted per term.
-
On-demand
$2.16 /GPU·hr
- Per GPU-second
- $0.000600
- Per 8 GPUs, per hour
- $17.28
- Per GPU-month, 730 h
- $1,576.80
Preview list price. It applies when your account opens.
-
Spot
From $0.87 /GPU·hr
Work that checkpoints and can wait. Follow the current spot price, or set a maximum hourly price. If the spot price rises above your max, the workload may stop.
"From" is the lowest spot price for that GPU. The spot price moves with supply and demand. The spot price always stays below the on-demand rate for the same GPU.
Spot price moves with supply and demand. Follow it, or set a max price per GPU-hour: you pay the spot price, and nodes stop if it rises above your max. A max price limits what you pay. It does not reserve capacity. To guarantee capacity, reserve it. How spot works
-
Reserved
Request a quote
Capacity you need on a date, guaranteed. 1 month, 6 months or 12+ months.
Ways to pay Prepaid credit, pay-as-you-go and Enterprise. Each one covers on-demand, spot and reserved capacity. How billing works
| Term | Commitment | Capacity | Rate |
|---|---|---|---|
| On-demandNone · Not held | None | Not held | $10.45 /GPU·hr |
| 1 monthTake-or-pay · Held for the term, on one fabric | Take-or-pay | Held for the term, on one fabric | Request a quote |
| 6 monthsTake-or-pay · Held for the term, on one fabric | Take-or-pay | Held for the term, on one fabric | Request a quote |
| 12+ monthsTake-or-pay · Held for the term, on one fabric | Take-or-pay | Held for the term, on one fabric | Request a quote |
| Term | Commitment | Capacity | Rate |
|---|---|---|---|
| On-demandNone · Not held | None | Not held | $9.35 /GPU·hr |
| 1 monthTake-or-pay · Held for the term, on one fabric | Take-or-pay | Held for the term, on one fabric | Request a quote |
| 6 monthsTake-or-pay · Held for the term, on one fabric | Take-or-pay | Held for the term, on one fabric | Request a quote |
| 12+ monthsTake-or-pay · Held for the term, on one fabric | Take-or-pay | Held for the term, on one fabric | Request a quote |
| Term | Commitment | Capacity | Rate |
|---|---|---|---|
| On-demandNone · Not held | None | Not held | $5.94 /GPU·hr |
| 1 monthTake-or-pay · Held for the term, on one fabric | Take-or-pay | Held for the term, on one fabric | Request a quote |
| 6 monthsTake-or-pay · Held for the term, on one fabric | Take-or-pay | Held for the term, on one fabric | Request a quote |
| 12+ monthsTake-or-pay · Held for the term, on one fabric | Take-or-pay | Held for the term, on one fabric | Request a quote |
| Term | Commitment | Capacity | Rate |
|---|---|---|---|
| On-demandNone · Not held | None | Not held | $5.40 /GPU·hr |
| 1 monthTake-or-pay · Held for the term, on one fabric | Take-or-pay | Held for the term, on one fabric | Request a quote |
| 6 monthsTake-or-pay · Held for the term, on one fabric | Take-or-pay | Held for the term, on one fabric | Request a quote |
| 12+ monthsTake-or-pay · Held for the term, on one fabric | Take-or-pay | Held for the term, on one fabric | Request a quote |
| Term | Commitment | Capacity | Rate |
|---|---|---|---|
| On-demandNone · Not held | None | Not held | $2.16 /GPU·hr |
| 1 monthTake-or-pay · Held for the term | Take-or-pay | Held for the term | Request a quote |
| 6 monthsTake-or-pay · Held for the term | Take-or-pay | Held for the term | Request a quote |
| 12+ monthsTake-or-pay · Held for the term | Take-or-pay | Held for the term | Request a quote |
| Term | Commitment | Capacity | Rate |
|---|---|---|---|
| On-demandNone · Not held | None | Not held | $1.86 /GPU·hr |
| 1 monthTake-or-pay · Held for the term | Take-or-pay | Held for the term | Request a quote |
| 6 monthsTake-or-pay · Held for the term | Take-or-pay | Held for the term | Request a quote |
| 12+ monthsTake-or-pay · Held for the term | Take-or-pay | Held for the term | Request a quote |
List price for the base shape: 1 GPU, 8 vCPU, 32 GiB. Larger shapes are quoted.
Reserved rates are quoted per term. Reservations are tied to a GPU type, a region and one InfiniBand fabric. Take-or-pay under a signed order form. No early cancellation without penalty (Terms §3).
Runs as single instances.
PCIe servers with 1 or 8 RTX PRO 6000 cards per machine. No NVLink and no InfiniBand.
RTX PRO 6000 is a PCIe card passed through whole to one machine: 1 GPU (24 vCPU · 218 GiB), or 8 gpus (192 vCPU · 1,744 GiB). No NVLink and no InfiniBand: these machines run as single instances. Connections: RTX PRO 6000 to 1 GPU; RTX PRO 6000 to 8 GPUs; 1 GPU to InfiniBand fabric; 8 GPUs to NVLink.
- memory
- 96 GB GDDR7 with ECC
- host
- PCIe 5.0 x16
- vCPU
- 24
- RAM
- 218 GiB
- network
- 400 Gb/s
1 GPU- vCPU
- 192
- RAM
- 1,744 GiB
- network
- 400 Gb/s
8 GPUsCards in one machine are not linkedto each other.
These machines do not join amulti-node cluster.
- memory
- 96 GB GDDR7 with ECC
- host
- PCIe 5.0 x16
- vCPU
- 24
- RAM
- 218 GiB
- network
- 400 Gb/s
1 GPU- vCPU
- 192
- RAM
- 1,744 GiB
- network
- 400 Gb/s
8 GPUsCards in one machine are not linked toeach other.
These machines do not join a multi-nodecluster.
- memory
- 96 GB GDDR7 with ECC
- host
- PCIe 5.0 x16
- vCPU
- 24
- RAM
- 218 GiB
- network
- 400 Gb/s
1 GPU- vCPU
- 192
- RAM
- 1,744 GiB
- network
- 400 Gb/s
8 GPUsCards in one machine are not linked to eachother.
These machines do not join a multi-nodecluster.
RTX PRO 6000 listing.
CPU, RAM and local NVMe vary by pool and are reported at placement. GPUs are passed through whole to one machine.
- Status
- Private Preview
- Runs as
- Single instances
- Capacity
- On-demand · Spot · Reserved
- Metering
- Metered by the second, billed hourly
| Parameter | Value | Scope |
|---|---|---|
| 1 GPU | 24 vCPU · 218 GiB | Per machine |
| 8 GPUs | 192 vCPU · 1,744 GiB | Per machine |
| InfiniBand | None | |
| Host network adapter | 400 Gb/s | Per machine |
NVIDIA RTX PRO 6000 Blackwell Server Edition on Fantasti: fantasti.ai/gpus/rtx-pro-6000 · Request capacity: fantasti.ai/contact
- Sheet
- 01 / 01
- Title
- RTX PRO 6000 datasheet
- Figures
- NVIDIA reference
- Reviewed
- 2026-10-10
Request RTX PRO 6000 capacity.
Tell us the count, the term and where it should run. We reply to a reviewed request, usually within one business day. A reserved term is quoted.