P40 vs RTX PRO 4000: Specs, VRAM & Price

Short answer: choose the RTX PRO 4000 when your workload uses FP8 quantized inference or serving. Choose the P40 when your workload is BF16/FP32-only and FP8 buys you nothing.

Both the P40 and the RTX PRO 4000 ship with 24 GB of VRAM.

You can rent both by the hour, no purchase required. P40 from $0.173/GPU/hr; RTX PRO 4000 from $0.419/GPU/hr. The P40 is 59% cheaper at the entry rate.

P40

Good for inference workloads on older-generation hardware

$0.173/hr
Lowest / GPU
$0.173/hr
Median / GPU
24 GB
VRAM
Pascal
Architecture

RTX PRO 4000

Good for inference and light fine-tuning workloads

$0.419/hr
Lowest / GPU
$0.560/hr
Median / GPU
24 GB
VRAM
Blackwell
Architecture

Last updated: 2026-10-10 20:17:43 UTCRefreshes hourly

Specs side by side

SpecP40RTX PRO 4000
VRAM24 GB24 GB
ArchitecturePascalBlackwell
Compute capability6.112.0
Precisions in hardwareFP32, FP16FP32, FP16, BF16, FP8, INT4
Form factorPCIe–
Regions12

On-demand price

RateP40RTX PRO 4000
Lowest per GPU/hr$0.173/GPU/hr (cheapest overall)$0.627/GPU/hr

Which is better for AI: P40 or RTX PRO 4000?

Choose the RTX PRO 4000 when your workload uses FP8 quantized inference or serving. Choose the P40 when your workload is BF16/FP32-only and FP8 buys you nothing.

P40 vs RTX PRO 4000: common questions

Is the P40 or the RTX PRO 4000 better for AI?

Short answer: choose the RTX PRO 4000 when your workload uses FP8 quantized inference or serving. Choose the P40 when your workload is BF16/FP32-only and FP8 buys you nothing.

Is the P40 or the RTX PRO 4000 cheaper to rent?

On Aquanode's live marketplace the P40 starts at $0.173/GPU/hr (median $0.173/GPU/hr) and the RTX PRO 4000 starts at $0.419/GPU/hr (median $0.560/GPU/hr). The P40 is the cheaper of the two at the entry rate, by 59%.

What is the difference between the P40 and the RTX PRO 4000?

Both cards report 24 GB of VRAM; the P40 is a Pascal part and the RTX PRO 4000 is Blackwell (compute capability 6.1 vs 12.0); only the RTX PRO 4000 supports BF16, FP8, INT4 compute. On price, the P40 lists from $0.173/GPU/hr and the RTX PRO 4000 from $0.419/GPU/hr.

How widely available are the P40 and the RTX PRO 4000?

The P40 is listed in 1 region and the RTX PRO 4000 in 2 regions right now.

How much does 1,000 GPU-hours cost on the P40 vs the RTX PRO 4000?

At the lowest rates listed today, 1,000 GPU-hours costs $173 on the P40 and $419 on the RTX PRO 4000, a difference of $246 for the same runtime. Rates are per GPU per hour and update hourly.

How much VRAM does the P40 have compared to the RTX PRO 4000?

Both the P40 and the RTX PRO 4000 ship with 24 GB of VRAM.

How this comparison is calculated

Every price is normalized to a per-GPU hourly rate using the same pipeline as every other pricing surface on Aquanode, and only the cheapest qualifying offer per provider is shown. A dash means that provider does not currently list that GPU.

Architecture, compute capability and supported precisions come from each vendor's published datasheet for that generation, not from the marketplace feed. This page regenerates at most once per hour.

Memory bandwidth, tensor TFLOPS, NVLink/interconnect and max board power are the same verified-from-datasheet figures shown on each model's own P40 page and RTX PRO 4000 page, which link the primary vendor source for that model. A dash here means Aquanode hasn't verified that figure from a primary source yet, never a guess.

Go further

Check what fits in each card's memory with the P40 VRAM calculator and the RTX PRO 4000 VRAM calculator, or browse every model by generation in the GPU index.

Submit the job. Everything after that is ours.

Sign up in 60 seconds. Pay for the GPU minutes you actually use.

© 2026 Aquanode. All rights reserved.

All trademarks, logos and brand names are the property of their respective owners.