Compare AI Models on Verified Benchmarks, Speed & Real Cost
Independent crowd preference Elo, SWE-bench coding tests, and live $/1M API pricing with zero synthetic bias.
Quality vs. Price Pareto Frontier
Top-Left = Best ValuePlotted by intelligence (Elo) vs cost (Out $). The teal line connects models with unbeatable price-to-performance.
Vertical ↑
Crowd vote (Elo)
Two hidden answers. A person picks the one they like. Higher Elo means more wins — not a school test or a coding exam.
Horizontal →
Output (answer) price / 1 million tokens (USD)
Cost to generate the answer. Provider list prices are USD
Live AI Model Leaderboard
Top 10 of 33 models.Top 30 of 33 models. Sortable rankings from dated verified benchmarks.
DeepSeek
3 more not shown
How Close is China's AI to US Flagships?
Quality is virtually tied on independent benchmarks, while open-weight Chinese frontiers offer substantial price-to-performance advantages.
Leaders Matchup
Frontier Trio Capability Radar
Comparing top Western standard models with China's leading frontier rival across 6 skill dimensions. Tap any spoke or dot to inspect.
Recent AI News & Benchmark Briefings
Independent analysis on model promotions, price reductions, and verified score movements.
Shipped & Updated
New and Refreshed Models
How to read this leaderboard
Plain-English methodology and leaderboard answers
- Preference Elo is a crowd vote from LMArena / Arena. People see two hidden answers and pick the one they like more. The model that wins more often gets a higher Elo. That means people preferred it — not that it passed a school test. It is not SWE-bench, not accuracy, and not a number we invent.
