Chutes: Models and Prices

Inference provider serving open-weight models.

Chutes served 41B tokens of other labs' open-weight models over Oct 3 to Oct 9, 2026 through a large public LLM router, #44 among providers.

Open-weight models

ModelFamilyShare of modelTokens, 7dInput $/MOutput $/MTokens/sUptime 3dContext
Qwen3.8 27BQwen2.8%16B$0.24$2.203599.8%262K
Kimi K2.6Kimi14.2%13B$0.50$2.854399.7%262K
Gemma 4 31BGemma1.7%7B$0.12$0.371078.0%131K
Qwen3.6 27BQwen87.7%3B$0.30$2.001692.2%262K
GLM 5.1GLM13.7%1B$0.98$3.082289.9%203K
Kimi K3Kimi<0.1%179M$3.00$15.003487.5%1.05M

List prices per million tokens as of October 9, 2026. Share: of all tokens served on that model; a dash where the volume is not reported.

Company

Website
chutes.ai
Headquarters
United States

Finance your GPUs

Serving tokens on your own GPUs? Tell us the cluster, the models you serve and the contracts behind them, and we will share it with partner funders that fit.

Prefer email? hello@amcompute.com