Hunyuan API Providers

Who serves Hunyuan, at what price and speed.

Hunyuan (Hy) is Tencent's model family. Tencent publishes open weights and serves almost all Hunyuan tokens itself through Tencent Cloud.

Tencent Cloud served 97% of the 10.1T Hunyuan tokens that went through a large public LLM router over Oct 3 to Oct 9, 2026. Tencent's own API served 97%; 7 providers list the family.

Who serves Hunyuan tokens

ProviderTokens, 7dShare
Tencent Cloud (lab)9.71T96.8%
NovitaAI168B1.7%
SiliconFlow152B1.5%
Phala650M<0.1%
AtlasCloud00.0%
inference.net00.0%
Reka AI00.0%

All Hunyuan models combined, Oct 3 to Oct 9, 2026.

Hy4 preview: prices and speed by provider

Output prices run from $2.50 to $2.50 per million tokens across 4 providers, median $2.50.

ProviderShareInput $/MOutput $/MCache read $/MTokens/sLatencyUptime 3dContext
Tencent Cloud96.0%$0.834$2.50$0.042434.46s99.8%1.05M
NovitaAI2.1%$0.834$2.50$0.042295.61s100.0%1M
SiliconFlow1.9%$0.834$2.50$0.042401.76s97.5%1.05M
DeepInfra-$0.834$2.50$0.0421211.59s97.8%1.05M

List prices as of October 9, 2026, cheapest endpoint per provider. Tokens/s and latency are medians over 30 minutes.

Hy3: prices and speed by provider

Output prices run from $0.528 to $0.80 per million tokens across 5 providers, median $0.58.

ProviderShareInput $/MOutput $/MCache read $/MTokens/sLatencyUptime 3dContext
Tencent Cloud100.0%$0.132$0.528$0.033364.38s100.0%262K
Phala<0.1%$0.15$0.64$0.04632.70s100.0%262K
NovitaAI-$0.14$0.58$0.035413.86s99.0%262K
GMICloud-$0.14$0.58$0.035662.62s99.8%262K
AtlasCloud-$0.20$0.80$0.05952.28s100.0%262K

List prices as of October 9, 2026, cheapest endpoint per provider. Tokens/s and latency are medians over 30 minutes.

All Hunyuan models

ModelTokens, 7dProvidersOutput $/MContextReleased
Hy4 preview7.96T4$2.501.05M2026-08-28
Hy32.10T5$0.528 to $0.80262K2026-07-06
Hy3 preview5B1$0.60262K2026-04-22
Hy-MT2-30B-A3B1B1$0.2958K2026-08-20
Hunyuan A13B Instruct50M1$0.57131K2025-07-08
Hy-MT2-7B47M1$0.2958K2026-08-19
Hy-MT2-1.8B46M1$0.1778K2026-08-20

Finance your GPUs

Serving Hunyuan on your own GPUs? Tell us the cluster, the models you serve and the contracts behind them, and we will share it with partner funders that fit.

Prefer email? hello@amcompute.com