B200 vs P40: Specs, VRAM & Price
Short answer: choose the B200 when 180 GB of VRAM is the gating requirement. Choose the P40 when 24 GB is enough for the workload and the lower hourly rate points that way.
The B200 has 180 GB of VRAM; the P40 has 24 GB.
You can rent both by the hour, no purchase required. B200 from $8.79/GPU/hr; P40 from $0.173/GPU/hr. The P40 is 98% cheaper at the entry rate.
B200
Good for training and serving frontier-scale LLMs
P40
Good for inference workloads on older-generation hardware
Last updated: 2026-10-10 20:28:46 UTCRefreshes hourly
Specs side by side
| Spec | B200 | P40 |
|---|---|---|
| VRAM | 180 GB | 24 GB |
| Memory bandwidth | 8 TB/s | – |
| Architecture | Blackwell | Pascal |
| Compute capability | 10.0 | 6.1 |
| Precisions in hardware | FP32, FP16, BF16, FP8, INT4 | FP32, FP16 |
| Peak FP16/BF16 tensor TFLOPS | 2,250 TFLOPS | – |
| Peak FP8 tensor TFLOPS | 4,500 TFLOPS | – |
| NVLink / multi-GPU interconnect | NVLink 5, 1.8 TB/s per GPU | – |
| Max board power (TDP) | Up to 1,000W | – |
| Form factor | SXM | PCIe |
| Regions | 3 | 1 |
On-demand price
| Rate | B200 | P40 |
|---|---|---|
| Lowest per GPU/hr | $8.79/GPU/hr | $0.173/GPU/hr (cheapest overall) |
Which is better for AI: B200 or P40?
Choose the B200 when 180 GB of VRAM is the gating requirement. Choose the P40 when 24 GB is enough for the workload and the lower hourly rate points that way.
B200 vs P40: common questions
Is the B200 or the P40 better for AI?
Short answer: choose the B200 when 180 GB of VRAM is the gating requirement. Choose the P40 when 24 GB is enough for the workload and the lower hourly rate points that way.
Is the B200 or the P40 cheaper to rent?
On Aquanode's live marketplace the B200 starts at $8.79/GPU/hr (median $8.79/GPU/hr) and the P40 starts at $0.173/GPU/hr (median $0.173/GPU/hr). The P40 is the cheaper of the two at the entry rate, by 98%.
What is the difference between the B200 and the P40?
The B200 has 180 GB of VRAM against the P40's 24 GB; the B200 is a Blackwell part and the P40 is Pascal (compute capability 10.0 vs 6.1); only the B200 supports BF16, FP8, INT4 compute. On price, the B200 lists from $8.79/GPU/hr and the P40 from $0.173/GPU/hr.
How widely available are the B200 and the P40?
The B200 is listed in 3 regions and the P40 in 1 region right now.
How much does 1,000 GPU-hours cost on the B200 vs the P40?
At the lowest rates listed today, 1,000 GPU-hours costs $8,789 on the B200 and $173 on the P40, a difference of $8,616 for the same runtime. Rates are per GPU per hour and update hourly.
How much VRAM does the B200 have compared to the P40?
The B200 has 180 GB of VRAM; the P40 has 24 GB.
How this comparison is calculated
Every price is normalized to a per-GPU hourly rate using the same pipeline as every other pricing surface on Aquanode, and only the cheapest qualifying offer per provider is shown. A dash means that provider does not currently list that GPU.
Architecture, compute capability and supported precisions come from each vendor's published datasheet for that generation, not from the marketplace feed. This page regenerates at most once per hour.
Memory bandwidth, tensor TFLOPS, NVLink/interconnect and max board power are the same verified-from-datasheet figures shown on each model's own B200 page and P40 page, which link the primary vendor source for that model. A dash here means Aquanode hasn't verified that figure from a primary source yet, never a guess.
Go further
Check what fits in each card's memory with the B200 VRAM calculator and the P40 VRAM calculator, see vendor-published throughput in the GPU benchmarks and specs table, or browse every model by generation in the GPU index.
More B200 comparisons
More P40 comparisons
Related reading: B200 pricing and specs, NVIDIA B200 guide, B300 vs B200, H200 vs B200 vs GB200, and The best GPUs for AI, ranked.