DeepInfra Company Profile

Serverless token API and dedicated deployments for open-weight models.

Profile

TypeInference platform
HeadquartersPalo Alto, CA
Websitedeepinfra.com

Inference

DeepInfra serves open-weight models through a large public LLM router. Rank is by tokens of other labs' open-weight models in the 7 days to Oct 9, 2026. DeepInfra models and prices.

ProviderOpen-weight modelsRank, 7 daysTop models
DeepInfra69#3DeepSeek V4.1 Flash, DeepSeek V4 Flash 0731, Mistral Nemo

Discuss a transaction

Looking to lend into GPU deals like these? Tell us what you finance and we will come back with the deals that fit.

Prefer email? hello@amcompute.com