Themed ranking
Fastest and most reliable
Ranked by real measured throughput (tokens/sec) at each model's best-performing provider, tracked daily.
As of September 23, 2026, OpenAI: gpt-oss-120b ranks #1 for fastest and most reliable at throughput (tok/s) 622 tok/s.
- #1 OpenAI: gpt-oss-120b - Throughput (tok/s) 622 tok/s
- #2 OpenAI: gpt-oss-20b (batch) - Throughput (tok/s) 453 tok/s
- #3 Google: Gemma 4 31B (free) - Throughput (tok/s) 304 tok/s
- #4 MiniMax: MiniMax M2.7 - Throughput (tok/s) 228 tok/s
- #5 Z.ai: GLM 5.2 (free) - Throughput (tok/s) 182 tok/s
- #6 Thinking Machines: Inkling (free) - Throughput (tok/s) 164 tok/s
- #7 Qwen: Qwen3.6 35B A3B - Throughput (tok/s) 158 tok/s
- #8 DeepSeek: DeepSeek V4.1 Flash (batch) - Throughput (tok/s) 154 tok/s
- #9 Google: Gemini 3.7 Flash (batch) - Throughput (tok/s) 144 tok/s
- #10 MiniMax: MiniMax M3 - Throughput (tok/s) 136 tok/s
Source: OpenRouter (openrouter.ai/rankings), as of September 23, 2026.
Benchmark scores: Artificial Analysis (artificialanalysis.ai) via OpenRouter (openrouter.ai/rankings).
Methodology: www.smophy.ai/benchmark/methodology
Token counts originate from each provider's own tokenizer and are not directly comparable across providers.
