L40SAda Lovelace
NVIDIA L40S
Inference, small-model fine-tuning, rendering and video on PCIe servers.
GDDR6 with ECC per GPU
48 GB
NVIDIA reference figure
- 1. PCIe x16 edge connector · PCIe Gen4
- 2. Passive heatsink · no fan on the card; the chassis moves the air
- 3. Cover · drawn cut away to show the fins
- 4. GPU, 48 GB GDDR6 with ECC · under the heatsink
- 5. Bracket · dual-slot, full height
- 6. 16-pin power connector · one per card
Point at a part or a row. The other one lights.
L40S specifications.
| Parameter | Value | Scope | Source |
|---|---|---|---|
| Memory | |||
| Memory | 48 GB GDDR6 with ECC | Per GPU | [1] |
| Memory bandwidth | 864 GB/s | Per GPU | [1] |
| Compute, peak · dense / sparse | |||
| FP4 Tensor Core | Not supported on Ada Lovelace | Per GPU | |
| FP8 Tensor Core | 733 / 1,466 TFLOPS | Per GPU | [1] |
| BF16 Tensor Core | 362 / 733 TFLOPS | Per GPU | [1] |
| Interconnect | |||
| NVLink | None | ||
| Host interface | PCIe Gen4 x16 · 64 GB/s bidirectional | Per GPU | [1] |
| Power and form factor | |||
| Power | 350 W | Per GPU | [1] |
| Form factor | PCIe card, dual-slot, full height, full length, passive cooling | [2] | |
Specifications are NVIDIA reference figures.
Best for
- Inference
- Fine-tuning smaller models
- 3D graphics, rendering and video (3 NVENC, 3 NVDEC, AV1)
Look elsewhere for
- Multi-node training that needs InfiniBand
Notes on the figures
- FP4: not supported on Ada Lovelace.
On-demand or reserved.
Metered by the second, billed hourly. Reserved rates are quoted per term.
-
On-demand
$1.86 /GPU·hr
- Per GPU-second
- $0.000517
- Per GPU-month, 730 h
- $1,357.80
List price for the base shape: 1 GPU, 8 vCPU, 32 GiB. Larger shapes are quoted.
-
Spot
Not offered
Spot is not offered on this listing.
-
Reserved
Request a quote
Capacity you need on a date, guaranteed. 1 month, 6 months or 12+ months.
Ways to pay Prepaid credit, pay-as-you-go and Enterprise. Each one covers on-demand, spot and reserved capacity. How billing works
| Term | Commitment | Capacity | Rate |
|---|---|---|---|
| On-demandNone · Not held | None | Not held | $10.45 /GPU·hr |
| 1 monthTake-or-pay · Held for the term, on one fabric | Take-or-pay | Held for the term, on one fabric | Request a quote |
| 6 monthsTake-or-pay · Held for the term, on one fabric | Take-or-pay | Held for the term, on one fabric | Request a quote |
| 12+ monthsTake-or-pay · Held for the term, on one fabric | Take-or-pay | Held for the term, on one fabric | Request a quote |
| Term | Commitment | Capacity | Rate |
|---|---|---|---|
| On-demandNone · Not held | None | Not held | $9.35 /GPU·hr |
| 1 monthTake-or-pay · Held for the term, on one fabric | Take-or-pay | Held for the term, on one fabric | Request a quote |
| 6 monthsTake-or-pay · Held for the term, on one fabric | Take-or-pay | Held for the term, on one fabric | Request a quote |
| 12+ monthsTake-or-pay · Held for the term, on one fabric | Take-or-pay | Held for the term, on one fabric | Request a quote |
| Term | Commitment | Capacity | Rate |
|---|---|---|---|
| On-demandNone · Not held | None | Not held | $5.94 /GPU·hr |
| 1 monthTake-or-pay · Held for the term, on one fabric | Take-or-pay | Held for the term, on one fabric | Request a quote |
| 6 monthsTake-or-pay · Held for the term, on one fabric | Take-or-pay | Held for the term, on one fabric | Request a quote |
| 12+ monthsTake-or-pay · Held for the term, on one fabric | Take-or-pay | Held for the term, on one fabric | Request a quote |
| Term | Commitment | Capacity | Rate |
|---|---|---|---|
| On-demandNone · Not held | None | Not held | $5.40 /GPU·hr |
| 1 monthTake-or-pay · Held for the term, on one fabric | Take-or-pay | Held for the term, on one fabric | Request a quote |
| 6 monthsTake-or-pay · Held for the term, on one fabric | Take-or-pay | Held for the term, on one fabric | Request a quote |
| 12+ monthsTake-or-pay · Held for the term, on one fabric | Take-or-pay | Held for the term, on one fabric | Request a quote |
| Term | Commitment | Capacity | Rate |
|---|---|---|---|
| On-demandNone · Not held | None | Not held | $2.16 /GPU·hr |
| 1 monthTake-or-pay · Held for the term | Take-or-pay | Held for the term | Request a quote |
| 6 monthsTake-or-pay · Held for the term | Take-or-pay | Held for the term | Request a quote |
| 12+ monthsTake-or-pay · Held for the term | Take-or-pay | Held for the term | Request a quote |
| Term | Commitment | Capacity | Rate |
|---|---|---|---|
| On-demandNone · Not held | None | Not held | $1.86 /GPU·hr |
| 1 monthTake-or-pay · Held for the term | Take-or-pay | Held for the term | Request a quote |
| 6 monthsTake-or-pay · Held for the term | Take-or-pay | Held for the term | Request a quote |
| 12+ monthsTake-or-pay · Held for the term | Take-or-pay | Held for the term | Request a quote |
List price for the base shape: 1 GPU, 8 vCPU, 32 GiB. Larger shapes are quoted.
Reserved rates are quoted per term. Reservations are tied to a GPU type, a region and one InfiniBand fabric. Take-or-pay under a signed order form. No early cancellation without penalty (Terms §3).
Runs as single instances.
PCIe servers with 1 to 4 L40S cards per machine. No NVLink and no InfiniBand.
L40S is a PCIe card passed through whole to one machine: 1 GPU (8 to 40 vCPU · 32 to 160 GiB), or up to 4 gpus (16 to 192 vCPU · 96 to 1,152 GiB). No NVLink and no InfiniBand: these machines run as single instances. Connections: L40S to 1 GPU; L40S to Up to 4 GPUs; 1 GPU to InfiniBand fabric; Up to 4 GPUs to NVLink.
- memory
- 48 GB GDDR6 with ECC
- host
- PCIe Gen4 x16
- vCPU
- 8 to 40
- RAM
- 32 to 160 GiB
1 GPU- vCPU
- 16 to 192
- RAM
- 96 to 1,152 GiB
Up to 4 GPUsCards in one machine are not linkedto each other.
These machines do not join amulti-node cluster.
- memory
- 48 GB GDDR6 with ECC
- host
- PCIe Gen4 x16
- vCPU
- 8 to 40
- RAM
- 32 to 160 GiB
1 GPU- vCPU
- 16 to 192
- RAM
- 96 to 1,152 GiB
Up to 4 GPUsCards in one machine are not linked toeach other.
These machines do not join a multi-nodecluster.
- memory
- 48 GB GDDR6 with ECC
- host
- PCIe Gen4 x16
- vCPU
- 8 to 40
- RAM
- 32 to 160 GiB
1 GPU- vCPU
- 16 to 192
- RAM
- 96 to 1,152 GiB
Up to 4 GPUsCards in one machine are not linked to eachother.
These machines do not join a multi-nodecluster.
L40S listing.
CPU, RAM and local NVMe vary by pool and are reported at placement. GPUs are passed through whole to one machine.
- Status
- Private Preview
- Runs as
- Single instances
- Capacity
- On-demand · Reserved
- Metering
- Metered by the second, billed hourly
| Parameter | Value | Scope |
|---|---|---|
| 1 GPU | 8 to 40 vCPU · 32 to 160 GiB | Per machine |
| Up to 4 GPUs | 16 to 192 vCPU · 96 to 1,152 GiB | Per machine |
| InfiniBand | None |
NVIDIA L40S on Fantasti: fantasti.ai/gpus/l40s · Request capacity: fantasti.ai/contact
- Sheet
- 01 / 01
- Title
- L40S datasheet
- Figures
- NVIDIA reference
- Reviewed
- 2026-10-10
Request L40S capacity.
Tell us the count, the term and where it should run. We reply to a reviewed request, usually within one business day. A reserved term is quoted.