Themed ranking
Best value
Ranked by quality per dollar - Artificial Analysis Intelligence Index divided by blended price per 1M tokens. A plain ratio, no invented weighting.
As of August 9, 2026, Ling-3.0-flash ranks #1 for best value at quality per $ 1200.0 pts/$.
- #1 Ling-3.0-flash - Quality per $ 1200.0 pts/$
- #2 inclusionAI: Ling-2.6-flash - Quality per $ 946.7 pts/$
- #3 Z.ai: GLM 5.2 (batch) - Quality per $ 489.3 pts/$
- #4 DeepSeek: DeepSeek V4 Flash 0423 - Quality per $ 460.4 pts/$
- #5 Tencent: Hy3 preview - Quality per $ 423.1 pts/$
- #6 OpenAI: gpt-oss-120b - Quality per $ 343.1 pts/$
- #7 OpenAI: gpt-oss-20b (free) - Quality per $ 276.4 pts/$
- #8 OpenAI: GPT-5.6 Luna (batch) - Quality per $ 232.4 pts/$
- #9 Xiaomi: MiMo-V2.5 - Quality per $ 217.1 pts/$
- #10 Qwen: Qwen3.5-9B - Quality per $ 193.8 pts/$
Top 15 by quality per dollar
This is a ratio, not a capability ranking: an ultra-cheap model with modest absolute quality can outscore a much stronger, pricier one, purely because price sits near zero. Each bar shows its own Quality and Price/1M alongside the ratio. The highlighted bar is today's single best value.
How this is calculated
Quality per $ = Intelligence Index ÷ blended price per 1M tokens - a plain ratio, not a capability ranking. An ultra-cheap model with modest absolute Quality can outrank a much stronger, pricier one purely because price sits near zero; the Quality and Price/1M columns above show the two real numbers behind every ratio.
What it can't tell you: consistency run-to-run, how costly a wrong answer is for your use case, or whether a task needs multiple attempts. For high-stakes work, weigh the absolute Quality score more than the ratio, and check measured endpoint uptime for consistency.
Source: OpenRouter (openrouter.ai/rankings), as of August 9, 2026.
Benchmark scores: Artificial Analysis (artificialanalysis.ai) via OpenRouter (openrouter.ai/rankings).
Methodology: smophy.ai/benchmark/methodology
Token counts originate from each provider's own tokenizer and are not directly comparable across providers.
