H200 vs P40: Specs, VRAM & Price

Short answer: choose the H200 when 141 GB of VRAM is the gating requirement. Choose the P40 when 24 GB is enough for the workload and the lower hourly rate points that way.

The H200 has 141 GB of VRAM; the P40 has 24 GB.

You can rent both by the hour, no purchase required. H200 from $3.95/GPU/hr; P40 from $0.173/GPU/hr. The P40 is 96% cheaper at the entry rate.

H200

Good for training and serving frontier-scale LLMs

$3.95/hr
Lowest / GPU
$5.82/hr
Median / GPU
141 GB
VRAM
Hopper
Architecture

P40

Good for inference workloads on older-generation hardware

$0.173/hr
Lowest / GPU
$0.173/hr
Median / GPU
24 GB
VRAM
Pascal
Architecture

Last updated: 2026-10-10 20:28:34 UTCRefreshes hourly

Specs side by side

SpecH200P40
VRAM141 GB24 GB
Memory bandwidth4.8 TB/s–
ArchitectureHopperPascal
Compute capability9.06.1
Precisions in hardwareFP32, FP16, BF16, FP8, INT4FP32, FP16
Peak FP16/BF16 tensor TFLOPS989 TFLOPS–
Peak FP8 tensor TFLOPS1,979 TFLOPS–
NVLink / multi-GPU interconnectNVLink 4, 900 GB/s bidirectional–
Max board power (TDP)Up to 700W (configurable)–
Form factorSXMPCIe
Regions41

On-demand price

RateH200P40
Lowest per GPU/hr$3.98/GPU/hr$0.173/GPU/hr (cheapest overall)

Which is better for AI: H200 or P40?

Choose the H200 when 141 GB of VRAM is the gating requirement. Choose the P40 when 24 GB is enough for the workload and the lower hourly rate points that way.

H200 vs P40: common questions

Is the H200 or the P40 better for AI?

Short answer: choose the H200 when 141 GB of VRAM is the gating requirement. Choose the P40 when 24 GB is enough for the workload and the lower hourly rate points that way.

Is the H200 or the P40 cheaper to rent?

On Aquanode's live marketplace the H200 starts at $3.95/GPU/hr (median $5.82/GPU/hr) and the P40 starts at $0.173/GPU/hr (median $0.173/GPU/hr). The P40 is the cheaper of the two at the entry rate, by 96%.

What is the difference between the H200 and the P40?

The H200 has 141 GB of VRAM against the P40's 24 GB; the H200 is a Hopper part and the P40 is Pascal (compute capability 9.0 vs 6.1); only the H200 supports BF16, FP8, INT4 compute. On price, the H200 lists from $3.95/GPU/hr and the P40 from $0.173/GPU/hr.

How widely available are the H200 and the P40?

The H200 is listed in 4 regions and the P40 in 1 region right now.

How much does 1,000 GPU-hours cost on the H200 vs the P40?

At the lowest rates listed today, 1,000 GPU-hours costs $3,949 on the H200 and $173 on the P40, a difference of $3,776 for the same runtime. Rates are per GPU per hour and update hourly.

How much VRAM does the H200 have compared to the P40?

The H200 has 141 GB of VRAM; the P40 has 24 GB.

How this comparison is calculated

Every price is normalized to a per-GPU hourly rate using the same pipeline as every other pricing surface on Aquanode, and only the cheapest qualifying offer per provider is shown. A dash means that provider does not currently list that GPU.

Architecture, compute capability and supported precisions come from each vendor's published datasheet for that generation, not from the marketplace feed. This page regenerates at most once per hour.

Memory bandwidth, tensor TFLOPS, NVLink/interconnect and max board power are the same verified-from-datasheet figures shown on each model's own H200 page and P40 page, which link the primary vendor source for that model. A dash here means Aquanode hasn't verified that figure from a primary source yet, never a guess.

Go further

Check what fits in each card's memory with the H200 VRAM calculator and the P40 VRAM calculator, see vendor-published throughput in the GPU benchmarks and specs table, or browse every model by generation in the GPU index.

Submit the job. Everything after that is ours.

Sign up in 60 seconds. Pay for the GPU minutes you actually use.

© 2026 Aquanode. All rights reserved.

All trademarks, logos and brand names are the property of their respective owners.