CoreWeave: Models and Prices

GPU cloud (Nasdaq: CRWV) that also serves open-weight models as a token API.

CoreWeave served 1.55T tokens of other labs' open-weight models over Oct 3 to Oct 9, 2026 through a large public LLM router, #11 among providers.

Open-weight models

ModelFamilyShare of modelTokens, 7dInput $/MOutput $/MTokens/sUptime 3dContext
DeepSeek V4.1 FlashDeepSeek1.2%451B$0.20$0.6517999.5%1.05M
DeepSeek V4 Flash 0731DeepSeek8.3%422B$0.13$0.28108100.0%262K
GLM 5.3 FlashGLM2.5%277B$0.15$0.5013799.7%1.05M
Gemma 4 31BGemma22.4%95B$0.10$0.3468100.0%262K
DeepSeek V4 Pro 0813DeepSeek6.3%67B$1.31$3.9611099.9%1.05M
gpt-oss-120bgpt-oss15.0%63B$0.03$0.174299.8%131K
gpt-oss-20bgpt-oss29.3%60B$0.03$0.13128100.0%131K
MiniMax M3MiniMax4.6%55B$0.23$0.966899.9%262K
GLM 5.2GLM3.4%52B$0.76$2.4216199.9%1.05M
Kimi K2.6Kimi7.2%7B$0.65$3.416499.8%262K
Gemma 4 26B A4B Gemma0.6%2B$0.10$0.306399.8%262K
Nemotron 3.5 LightningNemotron--$0.07$0.20235100.0%262K
Qwen3.8 27BQwen--$0.40$3.0011098.8%262K
Qwen3.6 35B A3BQwen--$0.25$1.2513699.9%262K
DeepSeek V3.1DeepSeek--$0.55$1.6557100.0%161K
Kimi K2.7 CodeKimi--$0.71$3.5027100.0%262K
Llama 3.1 8B InstructLlama--$0.22$0.2214399.9%131K
Llama 3.3 70B InstructLlama--$0.71$0.716399.7%128K
Granite 4.2 8BOther open-weight--$0.10$0.1582100.0%131K

List prices per million tokens as of October 9, 2026. Share: of all tokens served on that model; a dash where the volume is not reported.

Company

Headquarters
Livingston, NJ
Compute
Operates its own GPU fleet across leased and owned data centers.

Sources: CoreWeave.

Finance your GPUs

Serving tokens on your own GPUs? Tell us the cluster, the models you serve and the contracts behind them, and we will share it with partner funders that fit.

Prefer email? hello@amcompute.com