T4 GPU rental price
NVIDIA T4 · 16 GB VRAM. Live per-GPU pricing aggregated across 1 provider and 1 region.
The T4 (16GB GDDR6) currently rents from $0.158 per GPU per hour on Aquanode, across 1 provider.
T4 specs
T4 full specs
| Architecture | NVIDIA, launched 2018 |
| VRAM | 16GB GDDR6 |
| Memory bandwidth | 320+ GB/s |
| FP16 / BF16 tensor throughput | 65 TFLOPS (peak, dense) |
| TDP | 70W |
| Form factor | PCIe, low-profile, single-slot |
Specs sourced from the vendor's public datasheet. See the source.
What fits in 16GB GDDR6 of VRAM
| Model | Precision | Fits? |
|---|---|---|
| Llama 3.2 3B | FP16 | roughly 6GB. Fits comfortably. |
| Llama 3 8B | INT4 | ~5-6GB. Fits, with room for a modest batch. |
| Mistral 7B | FP16 | ~14GB. Fits, but with almost no KV-cache headroom. |
| Llama 3.1 70B | INT4 | ~35-40GB. Does not fit on a single 16GB card. |
Approximate, based on published parameter counts and standard bytes-per-parameter rules of thumb (FP16 ≈ 2 bytes/param, INT4 ≈ 0.5-0.6 bytes/param). Real footprint also depends on KV-cache size and framework overhead.
Good for
A 70W single-slot card that runs INT8/INT4 quantized inference cheaply. It is the first generation whose compute capability (7.5) clears the AWQ/GPTQ kernel floor, so small quantized models genuinely run on it. Useful when the job is high-volume small-model serving or video/AI pipelines and the hourly rate matters more than latency.
Not good for
It is a 2018 part: no BF16 and no FP8 at all, 16GB of VRAM, and 320+ GB/s of bandwidth. Anything trained in BF16 has to be converted, most modern serving stacks assume BF16 or FP8, and nothing above ~13B fits even quantized.
Get notified when the price drops
GPU supply moves hourly. Tell us what you're waiting for and we'll email you when a matching offer appears across any provider we track.
How this price is calculated
All prices on this page are normalized to a per-GPU hourly rate using each offer's authoritative GPU count, so that raw price is divided by the number of GPUs it actually covers; Akash reports its price as already per-GPU, so it is used as-listed. An offer with a missing, zero, or invalid GPU count is excluded entirely rather than published at a guessed rate.
No offers were excluded from this snapshot for a missing or invalid price. No offers were dropped as price outliers in this snapshot.
Only the cheapest qualifying offer per provider is shown in the table above. This page regenerates at most once per hour.
Frequently asked questions
How much does it cost to rent a T4?
Live T4 rental prices currently range from $0.158 to $0.158 per GPU per hour across 1 provider, with a median of $0.158 per GPU per hour.
Which provider has the cheapest T4?
Akash currently offers the lowest T4 rate on Aquanode's marketplace at $0.158 per GPU per hour in LIS.
How is the T4 price calculated?
All prices are normalized to a per-GPU hourly rate using each offer's authoritative GPU count, which the raw price is divided by; Akash reports its price as already per-GPU. Offers whose price can't be safely normalized, or whose rate is an extreme outlier against the rest of the market, are excluded.
How much does an T4 cost per hour?
Live T4 rental rates on Aquanode currently range from $0.158 to $0.158 per GPU per hour, with a median of $0.158/hr. See the live table above for current per-provider pricing.
How much VRAM does an T4 have?
The T4 has 16GB GDDR6, with 320+ GB/s of peak memory bandwidth.
Is renting cheaper than buying?
Renting avoids the upfront hardware cost and lets you match spend to actual usage. A rented T4 at $0.158/hr only costs money while it's running, whereas buying ties up capital in hardware that keeps depreciating whether it's in use or not. Which is cheaper depends on how continuously you'd run it; short or bursty workloads usually favor renting.