L40S GPU rental price
NVIDIA L40S · 48 GB VRAM. Live per-GPU pricing aggregated across 4 providers and 4 regions.
The L40S (48GB GDDR6 with ECC) currently rents from $0.790 per GPU per hour on Aquanode, across 4 providers.
The cheapest L40S offer right now (RunPod, $0.790/GPU/hr) is about 28% below the market median of $1.09/GPU/hr.
L40S specs
L40S full specs
| Architecture | NVIDIA, launched 2023 |
| VRAM | 48GB GDDR6 with ECC |
| Memory bandwidth | 864 GB/s |
| FP16 / BF16 tensor throughput | 362 TFLOPS (peak, dense) |
| FP8 tensor throughput | 733 TFLOPS (peak, dense) |
| TDP | 350W |
| Form factor | PCIe, dual-slot, air-cooled |
Specs sourced from the vendor's public datasheet. See the source.
What fits in 48GB GDDR6 with ECC of VRAM
| Model | Precision | Fits? |
|---|---|---|
| Llama 3 8B | FP16 | ~16GB. Fits with room for a large batch and KV cache. |
| Llama 3.1 70B | FP16 | ~140GB. Does not fit; needs multiple cards. |
| Llama 3.1 70B | INT4 | ~35-40GB. Fits on a single card. |
| Mistral 7B | FP8 | ~8GB. Fits comfortably, leaves room for high-concurrency serving. |
Approximate, based on published parameter counts and standard bytes-per-parameter rules of thumb (FP16 ≈ 2 bytes/param, INT4 ≈ 0.5-0.6 bytes/param). Real footprint also depends on KV-cache size and framework overhead.
Good for
A PCIe, air-cooled inference and fine-tuning card that doesn't need SXM/NVLink server infrastructure. 48GB is enough for most 7-13B models in FP16 and larger models in quantized form, and its FP8 tensor cores make it a solid throughput-per-dollar choice for serving.
Not good for
GDDR6 bandwidth (864 GB/s) is well below HBM parts like H100 (3.35 TB/s), so it's memory-bandwidth-bound on large-batch or long-context serving well before it's compute-bound, and no NVLink means no fast GPU-to-GPU path for multi-card training.
Get notified when the price drops
GPU supply moves hourly. Tell us what you're waiting for and we'll email you when a matching offer appears across any provider we track.
How this price is calculated
All prices on this page are normalized to a per-GPU hourly rate using each offer's authoritative GPU count, so that raw price is divided by the number of GPUs it actually covers; Akash reports its price as already per-GPU, so it is used as-listed. An offer with a missing, zero, or invalid GPU count is excluded entirely rather than published at a guessed rate.
No offers were excluded from this snapshot for a missing or invalid price. No offers were dropped as price outliers in this snapshot.
Only the cheapest qualifying offer per provider is shown in the table above. This page regenerates at most once per hour.
Frequently asked questions
How much does it cost to rent a L40S?
Live L40S rental prices currently range from $0.790 to $2.28 per GPU per hour across 4 providers, with a median of $1.09 per GPU per hour.
Which provider has the cheapest L40S?
RunPod currently offers the lowest L40S rate on Aquanode's marketplace at $0.790 per GPU per hour in Unknown.
How is the L40S price calculated?
All prices are normalized to a per-GPU hourly rate using each offer's authoritative GPU count, which the raw price is divided by; Akash reports its price as already per-GPU. Offers whose price can't be safely normalized, or whose rate is an extreme outlier against the rest of the market, are excluded.
How much does an L40S cost per hour?
Live L40S rental rates on Aquanode currently range from $0.790 to $2.28 per GPU per hour, with a median of $1.09/hr. See the live table above for current per-provider pricing.
How much VRAM does an L40S have?
The L40S has 48GB GDDR6 with ECC, with 864 GB/s of peak memory bandwidth.
Is renting cheaper than buying?
Renting avoids the upfront hardware cost and lets you match spend to actual usage. A rented L40S at $0.790/hr only costs money while it's running, whereas buying ties up capital in hardware that keeps depreciating whether it's in use or not. Which is cheaper depends on how continuously you'd run it; short or bursty workloads usually favor renting.