Hy3 models

3 Hy3 models on Hugging Face, from 298.8B to 298.8B parameters. At the precision each one is published in, the smallest needs about 334 GB of VRAM (Hy3-FP8, cheapest live fit: RTX 4090) and the largest about 668 GB (Hy3, cheapest live fit: RTX PRO 6000 WS). The cheapest way to run Hy3-FP8 is $3.09/hr.

Hy3 models

ModelParametersPublished asVRAM neededLive GPU fitEst. $/hr
Hy3298.8BBF16668 GBRTX PRO 6000 WS × 7$10.48/hr
Hy3-preview298.8BBF16668 GBRTX PRO 6000 WS × 7$10.48/hr
Hy3-FP8298.8BF8_E4M3334 GBRTX 4090 × 7$3.09/hr

VRAM is for the precision the model is published in, with the same overhead and an 8,192-token context assumed on every page; see the methodology. The fit shown is the lowest-priced single GPU type that holds the model at that precision, or the lowest-priced multi-GPU set (up to 8) when none does.

More from Tencent

Submit the job. Everything after that is ours.

Sign up in 60 seconds. Pay for the GPU minutes you actually use.

© 2026 Aquanode. All rights reserved.

All trademarks, logos and brand names are the property of their respective owners.