SmophyAI

Published daily · Next update 02:00 UTC

Provider reliability · tracked daily

Which provider is most reliable for each model?

Uptime and latency, per provider, per model - measured directly from OpenRouter's own endpoint data. The uptime figure is OpenRouter's own trailing-30-minute metric, captured once daily by our sync (not a live/real-time feed) - a daily snapshot of that window, not this instant. A model's benchmark score doesn't tell you whether you can actually depend on getting a response; this does.

Ranked by real usage, 30 models with provider data
#ModelProvidersBest providerUptime
01DeepSeek: DeepSeek V4 Flash 042320Cloudflare100.0%
02Xiaomi: MiMo-V2.56Novita98.9%
03Tencent: Hy36Baidu100.0%
04Z.ai: GLM 5.2 (batch)27Inceptron100.0%
05DeepSeek: DeepSeek V4 Pro18Together100.0%
06MiniMax: MiniMax M3 (batch)9Venice100.0%
07NVIDIA: Nemotron 3 Ultra (free)4Together99.7%
08StepFun: Step 3.7 Flash3DeepInfra99.4%
09Anthropic: Claude Opus 4.8 (batch)4Anthropic99.4%
10Anthropic: Claude Opus 4.7 (batch)4Amazon Bedrock100.0%
11OpenAI: GPT-5.6 Luna (batch)3Azure100.0%
12Anthropic: Claude Sonnet 5 (batch)4Anthropic100.0%
13Google: Gemini 3 Flash Preview (batch)2Google99.3%
14Anthropic: Claude Sonnet 4.6 (batch)4Google100.0%
15MoonshotAI: Kimi K310Moonshot AI100.0%
16Google: Gemini 2.5 Flash Lite (batch)2Google100.0%
17Google: Gemini 2.5 Flash (batch)2Google AI Studio99.9%
18Xiaomi: MiMo-V2.5-Pro7DigitalOcean100.0%
19Google: Gemini 3.1 Flash Lite (batch)2Google AI Studio99.0%
20Poolside: Laguna S 2.1 (free)1Poolside100.0%
21OpenAI: gpt-oss-120b18Cerebras100.0%
22Google: Gemini 3.6 Flash (batch)2Google99.9%
23Claude Opus 5 (batch)5Amazon Bedrock99.9%
24OpenAI: GPT-5.6 Sol (batch)3Azure100.0%
25DeepSeek: DeepSeek V3.214Google100.0%
26Google: Gemma 4 31B (free)16Cerebras100.0%
27Google: Gemma 4 26B A4B (free)9DeepInfra100.0%
28SpaceXAI: Grok 4.51xAI100.0%
29NVIDIA: Nemotron 3 Super (free)3Nebius100.0%
30Anthropic: Claude Fable 5 (batch)4Anthropic87.8%

Which host is most reliable overall?

Alibaba leads among hosts serving 3+ tracked models, at 100.0% median uptime across 5 models. Inverted from the table above: instead of "which host is best for this model," this ranks each host by measured reliability across every model it serves.

55 hosting providers, by median uptime
#ProviderModels servedMedian uptimeAvg uptimePeak throughput
01Alibaba5100.0%99.5%73 tok/s
02Cloudflare4100.0%100.0%77 tok/s
03Cerebras2100.0%100.0%777 tok/s
04Mara2100.0%100.0%246 tok/s
05Poolside2100.0%100.0%156 tok/s
06Inceptron1100.0%100.0%45 tok/s
07Sail Research1100.0%100.0%31 tok/s
08xAI1100.0%100.0%50 tok/s
09Azure13100.0%97.5%69 tok/s
10DeepSeek2100.0%100.0%80 tok/s
11Groq2100.0%100.0%367 tok/s
12Moonshot AI1100.0%100.0%27 tok/s
13Friendli3100.0%100.0%115 tok/s
14ModelRun1100.0%100.0%129 tok/s
15Tencent1100.0%100.0%46 tok/s
16Anthropic7100.0%97.9%69 tok/s
17Io Net1100.0%100.0%11 tok/s
18Decart1100.0%100.0%60 tok/s
19Amazon Bedrock1199.9%97.7%150 tok/s
20Baidu599.9%99.7%80 tok/s
21Novita1599.9%98.3%78 tok/s
22Ionstream299.8%99.8%25 tok/s
23Ambient299.8%99.8%88 tok/s
24AkashML299.8%99.8%57 tok/s
25Z.AI199.8%99.8%49 tok/s
26AtlasCloud899.7%99.6%49 tok/s
27SiliconFlow799.5%98.2%65 tok/s
28Venice999.5%94.3%83 tok/s
29DeepInfra1799.4%98.4%137 tok/s
30Crusoe299.4%99.4%99 tok/s
31BaseTen599.3%96.4%189 tok/s
32Google1599.3%97.1%120 tok/s
33Modal199.3%99.3%60 tok/s
34DigitalOcean899.3%98.8%106 tok/s
35Parasail999.2%98.6%123 tok/s
36GMICloud999.2%92.5%88 tok/s
37Fireworks599.2%97.4%147 tok/s
38Xiaomi299.2%99.2%47 tok/s
39Google AI Studio599.0%98.8%139 tok/s
40OpenAI698.9%98.7%154 tok/s
41CoreWeave598.8%98.2%106 tok/s
42Chutes298.7%98.7%25 tok/s
43Wafer298.6%98.6%43 tok/s
44Together798.6%89.9%128 tok/s
45StreamLake598.5%97.9%35 tok/s
46StepFun198.1%98.1%47 tok/s
47Morph598.0%97.4%35 tok/s
48Minimax297.7%97.7%59 tok/s
49Nebius297.7%97.7%183 tok/s
50SambaNova497.6%97.6%359 tok/s
51Mancer 2296.8%96.8%41 tok/s
52OpenInference295.2%95.2%14 tok/s
53NextBit194.3%94.3%40 tok/s
54Phala589.5%67.2%88 tok/s
55Claude Platform on AWS1---