DeepInfra Company Profile
Serverless token API and dedicated deployments for open-weight models.
Profile
| Type | Inference platform |
| Headquarters | Palo Alto, CA |
| Website | deepinfra.com |
Inference
DeepInfra serves open-weight models through a large public LLM router. Rank is by tokens of other labs' open-weight models in the 7 days to Oct 9, 2026. DeepInfra models and prices.
| Provider | Open-weight models | Rank, 7 days | Top models |
|---|---|---|---|
| DeepInfra | 69 | #3 | DeepSeek V4.1 Flash, DeepSeek V4 Flash 0731, Mistral Nemo |
Discuss a transaction
Looking to lend into GPU deals like these? Tell us what you finance and we will come back with the deals that fit.
Prefer email? hello@amcompute.com