Statistics · updated daily
AI model statistics, September 23, 2026
Anthropic: Claude Opus 5.5 (batch) leads on real quality benchmarks with an Artificial Analysis Intelligence Index of 57.6, among 94 models with published scores, as of September 23, 2026.
Every statistic below is computed directly from the same real data behind the rest of this site - never hand-written - so it updates automatically and never drifts. Free to cite with attribution; link directly to any numbered stat with its anchor.
- 01
Anthropic: Claude Opus 5.5 (batch) leads on real quality benchmarks with an Artificial Analysis Intelligence Index of 57.6, among 94 models with published scores, as of September 23, 2026.
- 02
Ox Alpha handles the largest real share of production traffic on OpenRouter, at 27.4% of tracked token volume, averaged over the trailing 80 days.
- 03
The median Intelligence Index across 94 benchmarked models is 23.5 - half of tracked models score higher, half lower.
- 04
inclusionAI: Ling-2.6-flash is the cheapest tracked model at $0.0150 per million blended tokens.
- 05
OpenAI: o1-pro is the most expensive tracked model at $262.50 per million blended tokens - a 17500x spread from the cheapest.
- 06
The median blended price across tracked models is $0.85 per million tokens.
- 07
inclusionAI: Ling 3.0 Flash delivers the best measured quality per dollar, at 654.0 intelligence-index points per dollar.
- 11
Real production usage is classified across 29distinct task categories, measured directly from OpenRouter's classified traffic.
- 12
DeepSeek: DeepSeek V4 Flash 0423 leads classification with 17% of classified production usage, as of September 23, 2026.
- 13
DeepSeek: DeepSeek V4 Flash 0423 leads data extraction with 16% of classified production usage, as of September 23, 2026.
- 14
Z.ai: GLM 5.3 Flash (batch) leads workflow execution with 12% of classified production usage, as of September 23, 2026.
- 15
DeepSeek: DeepSeek V4 Flash 0423 leads roleplay & fiction with 32% of classified production usage, as of September 23, 2026.
- 16
DeepSeek: DeepSeek V4 Flash 0423 leads content writing with 31% of classified production usage, as of September 23, 2026.
- 17
Tencent: Hy-MT2-1.8B leads translation with 19% of classified production usage, as of September 23, 2026.
- 18
DeepSeek: DeepSeek V4 Flash 0423 leads data transformation with 24% of classified production usage, as of September 23, 2026.
- 19
Z.ai: GLM 5.3 Flash (batch) leads code generation with 19% of classified production usage, as of September 23, 2026.
Real usage vs. real quality
Average usage share (trailing 80 days) plotted against Artificial Analysis quality score, for every model with both. No composite score, just two real numbers side by side - the gap between them is the whole story: Anthropic: Claude Opus 5.5 (batch) tops quality but ranks #339 by usage; Ox Alpha tops usage with a quality score of unpublished.
Hover, focus, or tap any point for the model name and exact values.
Source: OpenRouter (openrouter.ai/rankings), as of September 23, 2026.
Benchmark scores: Artificial Analysis (artificialanalysis.ai) via OpenRouter (openrouter.ai/rankings).
Methodology: www.smophy.ai/benchmark/methodology
Token counts originate from each provider's own tokenizer and are not directly comparable across providers.
