NVIDIA H100 GPU: Specs, VRAM, Price & Benchmarks (2026)

The NVIDIA H100 (80GB HBM3) launched in 2022. This guide covers its full specs, VRAM, SXM/PCIe differences where they apply, AI performance, and live per-GPU rental pricing on Aquanode.

Short answer on cost: the H100 rents from $2.19 per GPU per hour on Aquanode. Out of capacity right now.

How much VRAM does the H100 have?

The H100 has 80GB HBM3, with 3.35 TB/s of peak memory bandwidth.

  • VRAM: H100 80GB HBM3 vs A100 80GB HBM2e
  • VRAM: H100 80GB HBM3 vs H200 141GB HBM3e

What fits in 80GB HBM3 of VRAM

ModelPrecisionFits?
Llama 3 8B / similar 7-9B modelsFP16~16GB. Fits with plenty of headroom for a large KV cache.
Llama 3.1 70BFP16~140GB weights alone. Does not fit on one 80GB card; needs 2 GPUs.
Llama 3.1 70BINT4 (AWQ/GPTQ)~35-40GB. Fits comfortably on a single card, room for context.
Mixtral 8x7BFP16~94GB weights. Just over the 80GB limit; INT4 (~24GB) fits easily.

Approximate, based on published parameter counts and standard bytes-per-parameter rules of thumb (FP16 ≈ 2 bytes/param, INT4 ≈ 0.5-0.6 bytes/param). Real footprint also depends on KV-cache size and framework overhead.

What can the H100 run?

Popular open models from small to frontier scale, with the memory each needs and how many H100 cards (80GB HBM3 each) that takes.

ModelAs publishedFP8INT4
Qwen/Qwen3-8B 8.2BBF16: ~18.3 GB, 1 GPUFP8: ~9.2 GB, 1 GPUINT4: ~4.6 GB, 1 GPU
Qwen/Qwen2.5-14B-Instruct 14.8BBF16: ~33 GB, 1 GPUFP8: ~16.5 GB, 1 GPUINT4: ~8.3 GB, 1 GPU
Qwen/Qwen3-32B 32.8BBF16: ~73.2 GB, 1 GPUFP8: ~36.6 GB, 1 GPUINT4: ~18.3 GB, 1 GPU
Qwen/Qwen-72B 72.3BBF16: ~162 GB, 3 GPUsFP8: ~80.8 GB, 2 GPUsINT4: ~40.4 GB, 1 GPU
MiniMaxAI/MiniMax-M2.7 228.7BFP8: ~256 GB, 4 GPUs–INT4: ~128 GB, 2 GPUs
deepseek-ai/DeepSeek-R1 684.5BFP8: ~765 GB, 10 GPUs–INT4: ~383 GB, 5 GPUs

Estimates: weights at the stated precision plus a flat 20% for KV cache and overhead, at a moderate context length. A dash means the precision is not offered for that model (it is already published at that size). INT4 needs a published quantized checkpoint. Open any model for a per-GPU breakdown, or use the H100 VRAM calculator.

H100 VRAM calculator: check which models fit in its memory at each precision.

All models that fit in 80 GB: the open models whose weights and overhead fit, at native, FP8 and INT4 precision.

$2.19/GPU/hrOut of capacity right now
Lowest / GPU / hr
$4.39/GPU/hr
Median / GPU / hr
$4.39/GPU/hr
p90 / GPU / hr
100
Live offers
Last updated: 2026-10-10 21:01:23 UTCRefreshes hourly0 offer(s) excluded from this snapshot

H100 specs

H100 SXM5

VRAM80GB HBM3
Memory bandwidth3.35 TB/s
TDPUp to 700W (configurable)
Form factorSXM (DGX H100 / HGX H100 platform only)
InterconnectNVLink 4, 900 GB/s bidirectional

H100 PCIe

VRAM80GB HBM2e
Memory bandwidth~2 TB/s
TDP350W
Form factorStandard PCIe, dual-slot

Specs sourced from the vendor's public datasheet/product page. See the source.

Related reading: H100 vs H200, A100 vs H100, H100 and H200: SXM vs NVL vs PCIe, and The best GPUs for AI, ranked.

GPU Glossary: What is VRAM?, HBM, Tensor Cores, CUDA Cores, TFLOPS, NVLink vs PCIe

H100 SXM5 vs PCIe: which should you use?

  • Memory bandwidth: SXM5 delivers 3.35 TB/s vs ~2 TB/s on PCIe.
  • Interconnect: SXM5 has NVLink 4, 900 GB/s bidirectional. PCIe has no dedicated GPU-to-GPU interconnect.
  • Power and form factor: SXM5 is Up to 700W (configurable) in a SXM (DGX H100 / HGX H100 platform only) form factor. PCIe is 350W in a Standard PCIe, dual-slot form factor.

For large-scale distributed training, the variant with NVLink and the highest memory bandwidth is usually the right choice, since those advantages compound across a multi-GPU cluster. For single-GPU inference or fine-tuning, the lower-power variant often gives the same usable VRAM at a lower hourly rate.

H100 AI performance

The default choice for training and fine-tuning models up to the ~30-40B range on a single card, and for high-throughput FP8 inference. Its 1,979 TFLOPS dense FP8 figure and 900 GB/s NVLink make 8-GPU pretraining and serving runs practical. It's also the card with the deepest cloud/on-prem availability, so it's usually the cheapest per-FLOP hour to actually get.

  • Dense FP16/BF16 tensor throughput: 989 TFLOPS
  • Dense FP8 tensor throughput: 1,979 TFLOPS
  • Memory bandwidth: 3.35 TB/s

80GB isn't enough to hold a 70B-class model in FP16 on one card, and it has no native FP4 support, so it can't run the newest FP4-native model formats at full precision the way Blackwell parts can.

See how it stacks up against other cards in the GPU benchmarks and specs table.

The cheapest H100 offer right now ($2.19/GPU/hr) is about 50% below the market median of $4.39/GPU/hr. Out of capacity right now.

H100 price: what does it cost?

Buying. $25,000-$30,000 (PCIe card, standalone) (commonly quoted street price (no fixed retail; OEM channel)). A complete 8x SXM DGX H100 system runs roughly $250,000-$400,000. Secondary-market H100 pricing peaked around $80,000-$120,000 in 2023 and has moderated since, but standalone PCIe cards remain well above the figure above through most resellers.

Renting. The cheapest current on-demand rate for the H100 on Aquanode is $2.78/GPU/hr. Live rates range from $2.19 to $4.39 per GPU per hour, with a median of $4.39/GPU/hr. Billed by the offer's own terms; the table below shows every live rate. Out of capacity right now.

Region$/GPU/hrAvailableVRAMvCPURAM
–$2.19080 GB––
South Korea$2.77194 GB64336 GB
Canada$2.78180 GB28180 GB
Bkk, Th$2.98280 GB49862.5 GB
Des Moines, Us$3.00180 GB20128 GB
Noida, In$3.76880 GB24200 GB
Finland$4.953280 GB16200 GB
Virginia, Us$7.57080 GB16256 GB

H100 price history

Aquanode stores one snapshot of its GPU price index per UTC day. For the H100 that is 5 days so far, 2026-10-06 to 2026-10-10, so this is a short history, not a long-run trend. The lowest per-GPU rate was $2.19 on both 2026-10-06 and 2026-10-10.

Day (UTC)Lowest per GPU hourMedian per GPU hourData-center lowest per GPU hourOffers
2026-10-10$2.19$3.51$2.7874
2026-10-09$2.19$3.65$3.0074
2026-10-08$2.19$3.51$2.7871
2026-10-07$2.19$3.64$2.7872
2026-10-06$2.19$3.51$2.7873

Each row is the stored daily snapshot of the live index, copied as recorded. A dash means that day stored no figure. The same series is available as JSON and summarised in the monthly GPU price report.

How this price is calculated

All prices on this page are normalized to a per-GPU hourly rate using each offer's authoritative GPU count, so that raw price is divided by the number of GPUs it actually covers; some offers report price as already per-GPU, so those are used as-listed. An offer with a missing, zero, or invalid GPU count is excluded entirely rather than published at a guessed rate.

No offers were excluded from this snapshot for a missing or invalid price. No offers were dropped as price outliers in this snapshot.

Only the cheapest qualifying offer per provider is shown in the table above. This page regenerates at most once per hour.

H100 vs A100: how do they compare?

  • VRAM: H100 80GB HBM3 vs A100 80GB HBM2e
  • Memory bandwidth: 3.35 TB/s vs 2,039 GB/s
  • Dense FP16 tensor throughput: 989 TFLOPS vs 312 TFLOPS

On Aquanode right now, A100 starts at $0.936/GPU/hr against H100's $2.19/GPU/hr, about 57% less.

Full H100 vs A100 price comparison

H100 vs H200: how do they compare?

  • VRAM: H100 80GB HBM3 vs H200 141GB HBM3e
  • Memory bandwidth: 3.35 TB/s vs 4.8 TB/s
  • Dense FP8 tensor throughput: 1,979 TFLOPS vs 1,979 TFLOPS

On Aquanode right now, H100 starts at $2.19/GPU/hr against H200's $3.95/GPU/hr, about 45% less.

Full H100 vs H200 price comparison

Compare the H100 with other GPUs

Compare H100 with

Get notified when the price drops

GPU supply moves hourly. Tell us what you're waiting for and we'll email you when a matching offer appears across any provider we track.

One email per matching alert. Unsubscribe any time.

Rent a H100 on Aquanode

  • On-demand instances from $2.19/GPU/hr, billed by the provider's own terms, with no hardware procurement or long-term commitment.
  • 100 live offers across 8 regions today.
  • Set a price/availability alert above to hear the moment a cheaper or newly-available H100 offer appears.
  • Compare every H100 offer side by side, or browse the full multi-provider GPU marketplace.

Good for

The default choice for training and fine-tuning models up to the ~30-40B range on a single card, and for high-throughput FP8 inference. Its 1,979 TFLOPS dense FP8 figure and 900 GB/s NVLink make 8-GPU pretraining and serving runs practical. It's also the card with the deepest cloud/on-prem availability, so it's usually the cheapest per-FLOP hour to actually get.

Not good for

80GB isn't enough to hold a 70B-class model in FP16 on one card, and it has no native FP4 support, so it can't run the newest FP4-native model formats at full precision the way Blackwell parts can.

H100 FAQs

How much VRAM does the H100 have?

The H100 has 80GB HBM3, with 3.35 TB/s of peak memory bandwidth.

What is the NVIDIA H100?

The NVIDIA H100 is a GPU released in 2022, with 80GB HBM3 of memory and a Up to 700W (configurable) power envelope. See the full spec table above for interconnect, form factor and tensor-throughput details.

How much does it cost to rent a H100?

Live H100 rental prices currently range from $2.19 to $4.39 per GPU per hour, with a median of $4.39 per GPU per hour. Out of capacity right now.

What's the cheapest H100 rate?

The lowest current H100 rate on Aquanode is $2.19 per GPU per hour in . Out of capacity right now.

What is the difference between H100 SXM5 and PCIe?

SXM5 has 3.35 TB/s of memory bandwidth and NVLink 4, 900 GB/s bidirectional. PCIe has ~2 TB/s, no dedicated GPU-to-GPU interconnect. Both carry 80GB HBM3 / 80GB HBM2e respectively.

How does the H100 compare to the A100?

VRAM: H100 80GB HBM3 vs A100 80GB HBM2e Memory bandwidth: 3.35 TB/s vs 2,039 GB/s Dense FP16 tensor throughput: 989 TFLOPS vs 312 TFLOPS On Aquanode right now, A100 starts at $0.936/GPU/hr against H100's $2.19/GPU/hr, about 57% less. See the full H100 vs A100 comparison for a shared-provider price breakdown.

How does the H100 compare to the H200?

VRAM: H100 80GB HBM3 vs H200 141GB HBM3e Memory bandwidth: 3.35 TB/s vs 4.8 TB/s Dense FP8 tensor throughput: 1,979 TFLOPS vs 1,979 TFLOPS On Aquanode right now, H100 starts at $2.19/GPU/hr against H200's $3.95/GPU/hr, about 45% less. See the full H100 vs H200 comparison for a shared-provider price breakdown.

What is the H100 good for?

The default choice for training and fine-tuning models up to the ~30-40B range on a single card, and for high-throughput FP8 inference. Its 1,979 TFLOPS dense FP8 figure and 900 GB/s NVLink make 8-GPU pretraining and serving runs practical. It's also the card with the deepest cloud/on-prem availability, so it's usually the cheapest per-FLOP hour to actually get.

What are the H100's limitations?

80GB isn't enough to hold a 70B-class model in FP16 on one card, and it has no native FP4 support, so it can't run the newest FP4-native model formats at full precision the way Blackwell parts can.

Is renting cheaper than buying?

Renting avoids the upfront hardware cost and lets you match spend to actual usage. A rented H100 at $2.19/hr only costs money while it's running, whereas buying ties up capital in hardware that keeps depreciating whether it's in use or not. Buying outright runs $25,000-$30,000 (PCIe card, standalone) (commonly quoted street price (no fixed retail; OEM channel)). Which is cheaper depends on how continuously you'd run it; short or bursty workloads usually favor renting. Out of capacity right now.

How is the H100 price calculated?

All prices are normalized to a per-GPU hourly rate using each offer's authoritative GPU count, which the raw price is divided by; some offers report price as already per-GPU. Offers whose price can't be safely normalized, or whose rate is an extreme outlier against the rest of the market, are excluded.

H100 price by region

Related guides

Other models in the same generation, then the rest of the GPU index.

Submit the job. Everything after that is ours.

Sign up in 60 seconds. Pay for the GPU minutes you actually use.

© 2026 Aquanode. All rights reserved.

All trademarks, logos and brand names are the property of their respective owners.