Compare AI Models on Verified Benchmarks, Speed & Real Cost
Independent crowd preference Elo, SWE-bench coding tests, and live $/1M API pricing with zero synthetic bias.
Quality vs. Price Pareto Frontier
Top-Left = Best ValuePlotted by intelligence (Elo) vs cost (Out $). The teal line connects models with unbeatable price-to-performance.
Vertical ↑
Crowd vote (Elo)
Two hidden answers. A person picks the one they like. Higher Elo means more wins — not a school test or a coding exam.
Horizontal →
Output (answer) price / 1 million tokens (USD)
Cost to generate the answer. Provider list prices are USD
Live AI Model Leaderboard
Top 10 of 26 models.Top 26 of 26 models. Sortable rankings from dated verified benchmarks.
DeepSeek
DeepSeek
How Close is China's AI to US Flagships?
Comparing 🇺🇸 Claude Opus 5 (Anthropic) vs 🇨🇳 GLM-5.3 (Zhipu). Quality is virtually tied, while China is surpassing on price and open weights.
Leaders Matchup
Multi-Dimensional Capability Radar
Each spoke is a skill. Farther from the center is better (0–100th percentile). Tap any dot to inspect details.
Recent AI News & Benchmark Briefings
Independent analysis on model promotions, price reductions, and verified score movements.
Shipped & Updated
New and Refreshed Models
How to read this leaderboard
Plain-English methodology and leaderboard answers
- Preference Elo is a crowd vote from LMArena / Arena. People see two hidden answers and pick the one they like more. The model that wins more often gets a higher Elo. That means people preferred it — not that it passed a school test. It is not SWE-bench, not accuracy, and not a number we invent.
