Wafer Company Profile

Inference provider whose agents rewrite kernels, batching, scheduling and memory layout for open-weight models.

Profile

TypeInference platform
HeadquartersUnited States
Websitewafer.ai

Inference

Wafer serves open-weight models through a large public LLM router. Rank is by tokens of other labs' open-weight models in the 7 days to Oct 9, 2026. Wafer models and prices.

ProviderOpen-weight modelsRank, 7 daysTop models
Wafer10#7DeepSeek V4.1 Flash, GLM 5.3, GLM 5.2

Discuss a transaction

Looking to lend into GPU deals like these? Tell us what you finance and we will come back with the deals that fit.

Prefer email? hello@amcompute.com