B200 GPU rental price
NVIDIA B200 · 180 GB VRAM. Live per-GPU pricing aggregated across 3 providers and 3 regions.
The B200 (180GB HBM3e) currently rents from $6.79 per GPU per hour on Aquanode, across 3 providers.
B200 specs
B200 full specs
| Architecture | NVIDIA, launched 2024 |
| VRAM | 180GB HBM3e |
| Memory bandwidth | 8 TB/s |
| FP16 / BF16 tensor throughput | 2,250 TFLOPS (peak, dense) |
| FP8 tensor throughput | 4,500 TFLOPS (peak, dense) |
| Interconnect | NVLink 5, 1.8 TB/s per GPU |
| TDP | Up to 1,000W |
| Form factor | SXM |
Specs sourced from the vendor's public datasheet. See the source.
What fits in 180GB HBM3e of VRAM
| Model | Precision | Fits? |
|---|---|---|
| Llama 3.1 70B | FP16 | ~140GB. Fits on one card with headroom to spare. |
| Llama 3.1 405B | FP8 | ~410GB. Needs 3+ cards even at 180GB each. |
| Llama 3.1 405B | INT4 | ~200-230GB. Still needs 2 cards. |
| Mixtral 8x22B | FP8 | ~140GB. Fits on a single card. |
Approximate, based on published parameter counts and standard bytes-per-parameter rules of thumb (FP16 ≈ 2 bytes/param, INT4 ≈ 0.5-0.6 bytes/param). Real footprint also depends on KV-cache size and framework overhead.
Good for
The highest-throughput card on this marketplace: more than double H100's FP8 throughput and over double the memory bandwidth, plus native FP4 tensor-core support for the newest FP4-quantized model formats. Built for frontier-scale training and the highest-throughput inference deployments.
Not good for
Supply is thin and the hourly rate is the highest on the marketplace. It's the wrong card for a workload that fits comfortably on an H100 or H200, since the extra throughput goes unused and the cost premium doesn't pay for itself.
Get notified when the price drops
GPU supply moves hourly. Tell us what you're waiting for and we'll email you when a matching offer appears across any provider we track.
How this price is calculated
All prices on this page are normalized to a per-GPU hourly rate using each offer's authoritative GPU count, so that raw price is divided by the number of GPUs it actually covers; Akash reports its price as already per-GPU, so it is used as-listed. An offer with a missing, zero, or invalid GPU count is excluded entirely rather than published at a guessed rate.
No offers were excluded from this snapshot for a missing or invalid price. No offers were dropped as price outliers in this snapshot.
Only the cheapest qualifying offer per provider is shown in the table above. This page regenerates at most once per hour.
Frequently asked questions
How much does it cost to rent a B200?
Live B200 rental prices currently range from $6.79 to $9.44 per GPU per hour across 3 providers, with a median of $6.79 per GPU per hour.
Which provider has the cheapest B200?
RunPod currently offers the lowest B200 rate on Aquanode's marketplace at $6.79 per GPU per hour in Unknown.
How is the B200 price calculated?
All prices are normalized to a per-GPU hourly rate using each offer's authoritative GPU count, which the raw price is divided by; Akash reports its price as already per-GPU. Offers whose price can't be safely normalized, or whose rate is an extreme outlier against the rest of the market, are excluded.
How much does an B200 cost per hour?
Live B200 rental rates on Aquanode currently range from $6.79 to $9.44 per GPU per hour, with a median of $6.79/hr. See the live table above for current per-provider pricing.
How much VRAM does an B200 have?
The B200 has 180GB HBM3e, with 8 TB/s of peak memory bandwidth.
Is renting cheaper than buying?
Renting avoids the upfront hardware cost and lets you match spend to actual usage. A rented B200 at $6.79/hr only costs money while it's running, whereas buying ties up capital in hardware that keeps depreciating whether it's in use or not. Which is cheaper depends on how continuously you'd run it; short or bursty workloads usually favor renting.