L40 GPU rental price
NVIDIA L40 · 48 GB VRAM. Live per-GPU pricing aggregated across 3 providers and 3 regions.
The L40 (48GB GDDR6 with ECC) currently rents from $0.690 per GPU per hour on Aquanode, across 3 providers.
The cheapest L40 offer right now (RunPod, $0.690/GPU/hr) is about 16% below the market median of $0.820/GPU/hr.
L40 specs
L40 full specs
| Architecture | NVIDIA, launched 2022 |
| VRAM | 48GB GDDR6 with ECC |
| Memory bandwidth | 864 GB/s |
| FP16 / BF16 tensor throughput | 181.05 TFLOPS (peak, dense) |
| FP8 tensor throughput | 362 TFLOPS (peak, dense) |
| TDP | 300W |
| Form factor | PCIe, dual-slot, passive |
Specs sourced from the vendor's public datasheet. See the source.
What fits in 48GB GDDR6 with ECC of VRAM
| Model | Precision | Fits? |
|---|---|---|
| Llama 3 8B | FP16 | roughly 16GB. Fits with room for a real KV cache. |
| Mixtral 8x7B | INT4 | ~24GB. Fits on a single card. |
| Llama 3.1 70B | INT4 | ~35-40GB. Fits, but with little headroom for context. |
| Llama 3.1 70B | FP16 | ~140GB. Does not fit; needs 3+ cards. |
Approximate, based on published parameter counts and standard bytes-per-parameter rules of thumb (FP16 ≈ 2 bytes/param, INT4 ≈ 0.5-0.6 bytes/param). Real footprint also depends on KV-cache size and framework overhead.
Good for
The passively-cooled data-center sibling of the L40S: same 48GB of ECC GDDR6 and the same 864 GB/s, with FP8 tensor cores for quantized serving. It's a reasonable card for 7-13B inference and fine-tuning in a rack that can't take SXM parts, and the ECC matters if a long run can't tolerate a silent bit-flip.
Not good for
Roughly half the L40S's dense tensor throughput on the same memory system, so it's the slower card at the same VRAM, and with no NVLink, multi-card training falls back to PCIe. GDDR6 bandwidth also caps large-batch serving well before compute does.
Get notified when the price drops
GPU supply moves hourly. Tell us what you're waiting for and we'll email you when a matching offer appears across any provider we track.
How this price is calculated
All prices on this page are normalized to a per-GPU hourly rate using each offer's authoritative GPU count, so that raw price is divided by the number of GPUs it actually covers; Akash reports its price as already per-GPU, so it is used as-listed. An offer with a missing, zero, or invalid GPU count is excluded entirely rather than published at a guessed rate.
No offers were excluded from this snapshot for a missing or invalid price. No offers were dropped as price outliers in this snapshot.
Only the cheapest qualifying offer per provider is shown in the table above. This page regenerates at most once per hour.
Frequently asked questions
How much does it cost to rent a L40?
Live L40 rental prices currently range from $0.690 to $1.01 per GPU per hour across 3 providers, with a median of $0.820 per GPU per hour.
Which provider has the cheapest L40?
RunPod currently offers the lowest L40 rate on Aquanode's marketplace at $0.690 per GPU per hour in Unknown.
How is the L40 price calculated?
All prices are normalized to a per-GPU hourly rate using each offer's authoritative GPU count, which the raw price is divided by; Akash reports its price as already per-GPU. Offers whose price can't be safely normalized, or whose rate is an extreme outlier against the rest of the market, are excluded.
How much does an L40 cost per hour?
Live L40 rental rates on Aquanode currently range from $0.690 to $1.01 per GPU per hour, with a median of $0.820/hr. See the live table above for current per-provider pricing.
How much VRAM does an L40 have?
The L40 has 48GB GDDR6 with ECC, with 864 GB/s of peak memory bandwidth.
Is renting cheaper than buying?
Renting avoids the upfront hardware cost and lets you match spend to actual usage. A rented L40 at $0.690/hr only costs money while it's running, whereas buying ties up capital in hardware that keeps depreciating whether it's in use or not. Which is cheaper depends on how continuously you'd run it; short or bursty workloads usually favor renting.