NextBit: Models and Prices

EU-only inference API on bare-metal GPU clusters.

NextBit served 154B tokens of other labs' open-weight models over Oct 3 to Oct 9, 2026 through a large public LLM router, #35 among providers.

Open-weight models

ModelFamilyShare of modelTokens, 7dInput $/MOutput $/MTokens/sUptime 3dContext
Gemma 4 26B A4B Gemma27.9%74B$0.068$0.22584100.0%262K
Gemma 3 12BGemma100.0%2B$0.05$0.156996.3%131K
Gemma 3 27BGemma2.9%1B$0.08$0.304495.0%131K
Qwen3 14BQwen100.0%943M$0.10$0.225899.8%41K
Llama 3.3 Euryale 70BOther open-weight100.0%310M$0.65$0.751197.0%131K
Gemma 2 27BGemma100.0%28M$0.65$0.6542100.0%8K
ReMM SLERP 13BOther open-weight5.9%5M$0.45$0.6524100.0%6K

List prices per million tokens as of October 9, 2026. Share: of all tokens served on that model; a dash where the volume is not reported.

Company

Headquarters
Zaragoza, Spain
Compute
Says it owns and operates its own GPU infrastructure in Spain.

Sources: Nextbit on Product Hunt.

Finance your GPUs

Serving tokens on your own GPUs? Tell us the cluster, the models you serve and the contracts behind them, and we will share it with partner funders that fit.

Prefer email? hello@amcompute.com