Inference Uptime Tracker

Open-weight inference providers ranked by uptime, October 8 to 10, 2026.

Uptime by provider

5 of 51 providers held 99.9% uptime or better; 26 fell below 99%. Uptime is the share of measured minutes in which a provider's endpoints served requests through a large public LLM router. The router publishes only the last three days, so read this as a snapshot. It is one channel among many: direct and enterprise traffic is not included.

#ProviderUptimeToken-weightedDowntime a monthWorst model-dayTokens, 7 days
1Relace99.99%100.00%5 min99.8%, DeepSeek V4 Pro 0423, Oct 88.91T
2Moonshot AI99.97%100.00%13 min99.8%, Kimi K2.6, Oct 9312B
3inference.net99.96%99.98%18 min97.8%, Schematron V2 Turbo, Oct 94.09T
4Tencent Cloud99.95%99.90%21 min99.7%, Hy4 preview, Oct 89.71T
5Modal99.94%99.99%28 min97.7%, Qwen3.8 2.4T A95B, Oct 8460B
6MiniMax99.88%99.95%53 min96.2%, MiniMax M2.5, Oct 81.07T
7ModelRun [by Modular]99.83%99.91%1 h 14 min99.0%, Gemma 4 31B, Oct 8283B
8CoreWeave99.82%99.75%1 h 19 min97.2%, Qwen3.8 27B, Oct 91.55T
9Baseten99.81%99.69%1 h 22 min98.4%, gpt-oss-120b, Oct 10598B
10AkashML99.78%99.83%1 h 34 min98.7%, Qwen3.6 35B A3B, Oct 889B
10Amazon Bedrock99.78%99.34%1 h 37 min98.3%, Kimi K3, Oct 965B
12Wafer99.74%99.93%1 h 52 min96.3%, Kimi K3, Oct 82.57T
13Darkbloom99.64%99.96%2 h 36 min96.8%, Qwen3.8 27B, Oct 10108B
14Cloudflare99.62%99.56%2 h 44 min95.5%, GLM 5.2, Oct 1076B
15Z.ai99.59%100.00%2 h 59 min89.5%, GLM 4.5 Air, Oct 81.30T
16DekaLLM99.50%99.63%3 h 36 min96.8%, GLM 5.3 Flash, Oct 9163B
17Friendli99.49%99.19%3 h 40 min96.3%, Gemma 4 31B, Oct 8350B
18Mistral99.46%99.68%3 h 53 min96.5%, Ministral 3 14B 2512, Oct 8600B
19Decart99.44%99.86%4 h 1 min96.3%, Kimi K3, Oct 10973B
20Makora99.31%99.61%4 h 57 min96.5%, GLM 5.3, Oct 8414B
20Sail Research99.31%98.68%4 h 58 min92.4%, DeepSeek V4 Flash 0731, Oct 81.71T
22io.net99.24%99.65%5 h 30 min92.7%, Gemma 4 31B, Oct 897B
23Parasail99.19%99.90%5 h 51 min43.0%, Qwen3 VL 8B Instruct, Oct 102.67T
24Inceptron99.17%99.28%5 h 57 min94.9%, GLM 5.3 Flash, Oct 8160B
25Alibaba Cloud Int.99.11%99.60%6 h 22 min88.6%, Qwen3.5-122B-A10B, Oct 101.01T
26DigitalOcean98.90%99.74%7 h 55 min68.6%, Qwen3.5 397B A17B, Oct 101.37T
27Crusoe98.51%99.29%10 h 46 min95.3%, gpt-oss-120b, Oct 9202B
28DeepInfra98.49%99.97%10 h 50 min70.1%, DeepSeek V3, Oct 86.43T
29Baidu Qianfan98.22%98.48%12 h 48 min79.1%, DeepSeek V4 Pro 0423, Oct 8500B
30Xiaomi98.12%99.54%13 h 32 min93.1%, MiMo-V2.5, Oct 1011.8T
31Reka AI98.08%98.97%13 h 51 min87.3%, DeepSeek V4 Pro 0423, Oct 8279B
32Ionstream98.02%97.92%14 h 17 min90.3%, DeepSeek V4 Pro 0813, Oct 8103B
32NextBit98.02%99.83%14 h 17 min92.9%, Gemma 3 27B, Oct 978B
34Morph97.57%97.97%17 h 31 min90.3%, GLM 5.3 Flash, Oct 81.02T
35Phala97.52%98.44%17 h 51 min70.3%, Qwen3.8 27B, Oct 10311B
36Together97.51%99.69%17 h 54 min82.2%, Inkling, Oct 811.5T
37AtlasCloud97.10%99.37%20 h 54 min49.5%, Qwen3.5-122B-A10B, Oct 101.17T
38Open Inference97.05%96.46%21 h 13 min82.2%, DeepSeek V4.1 Flash, Oct 10759B
39Mancer97.04%98.23%21 h 20 min70.1%, GLM 4.7, Oct 917B
40Venice96.92%99.22%22 h 12 min69.8%, Qwen3 Coder 480B A35B, Oct 8465B
41SiliconFlow96.86%97.59%22 h 35 min68.7%, gpt-oss-120b, Oct 9893B
42GMICloud96.77%97.30%23 h 16 min64.5%, MiMo-V2.6-Flash, Oct 101.45T
43MARA96.49%96.99%25 h 18 min55.0%, DeepSeek V3.2, Oct 85B
44Fireworks95.81%89.81%30 h 12 min39.9%, DeepSeek V4.1 Flash, Oct 102.20T
45StreamLake95.59%99.73%31 h 45 min8.2%, MiMo-V2.5, Oct 102.52T
46NovitaAI95.47%98.29%32 h 36 min5.4%, GLM 4.7 Flash, Oct 103.04T
47Nebius Token Factory94.93%93.43%36 h 30 min74.8%, GLM 5.2, Oct 834B
47SambaNova94.93%95.98%36 h 30 min85.0%, Gemma 4 31B, Oct 108B
49Groq93.60%92.12%46 h 4 min2.6%, gpt-oss-20b, Oct 10161B
50Chutes93.24%95.53%48 h 40 min68.3%, Gemma 4 31B, Oct 841B
51Google Vertex90.42%-69 h 1 min3.0%, gpt-oss-120b, Oct 9-

Uptime is weighted by measured minutes; token-weighted weights each model by the provider's tokens on it. Downtime a month applies the three-day rate to 30 days. Ties share a rank.

Discuss a transaction

Serving open-weight models on your own GPUs, or lending to a provider that does? Tell us about the cluster and the financing you need.

Prefer email? hello@amcompute.com