Independent crowd preference Elo, SWE-bench coding tests, and live $/1M API pricing with zero synthetic bias.
Plotted by intelligence (Elo) vs cost (Out $). The teal line connects models with unbeatable price-to-performance.
Vertical ↑
Crowd vote (Elo)
Two hidden answers. A person picks the one they like. Higher Elo means more wins — not a school test or a coding exam.
Horizontal →
Output (answer) price / 1 million tokens (USD)
Cost to generate the answer. Provider list prices are USD
Top 4 of 4 models.Top 4 of 4 models. Rankings from blind preference votes, generation latency, and price per 1k images.
Black Forest Labs
Black Forest Labs
Alibaba
Ideogram
Quality is virtually tied on independent benchmarks, while open-weight Chinese frontiers offer substantial price-to-performance advantages.
GLM-5.3 matches 96% of Claude Opus 5's intelligence at 96% lower cost.
Comparing top Western standard models with China's leading frontier rival across 6 skill dimensions. Tap any spoke or dot to inspect.
Independent analysis on model promotions, price reductions, and verified score movements.
Shipped & Updated
Plain-English methodology and leaderboard answers
Top flagship duel for #1 general intelligence
Western standard vs Chinese frontier price/perf leader
Top self-hostable open-weight models compared
Fast reasoning & latency vs low $/1M API cost
Gemini 3.7 Flash
Google Aug 13 2026 workhorse. Official intro price $0.75/$3.75 per 1M through Dec 31 2026; 1,048,576-token context (Google blog).