CompareLLM
CompareLLM.ai
Live
Explore
Compare
Decide
Track
Learn
…
CompareLLM
CompareLLM.ai
Precision Benchmarks

Programmatic, dated AI model benchmarks, head-to-head comparisons, and Stack Engine presets.

Daily ingest · 06:00 UTC

Analytics & Benchmarks

  • AI Model Leaderboard
  • Head-to-Head Compare Hub
  • Models Directory
  • Stack Engine Presets
  • Frontier Models
  • Open Weights Catalog

Guides & Intent Lists

  • Best LLM Lists (2026)
  • Best Coding LLM
  • Best Cheap LLM
  • Fastest Low-Latency LLM
  • Claude vs GPT Benchmark
  • What is Elo?
  • Methodology Guides
  • News & Dispatches

Transparency & API

  • Evaluation Methodology
  • Benchmark Changelog
  • Public JSON API
  • llms.txt Specification
  • Privacy Policy
  • Sign In / Account

© 2026 CompareLLM. Public benchmark data aggregated from Arena Elo, LiveBench, SWE-bench & OpenRouter.

Every score has a dated snapshot.

Theme:
Currency:
  1. Home
  2. What Changed
AI Model Intelligence Hub
Live Audit Stream

What Changed in AI Today?

Every meaningful shift in the AI model landscape: benchmark adjustments, price cuts, speed milestones, and new model releases converted into actionable decision intelligence.

Check your followed models|Find a replacement model
Showing 20 of 508 verified market changes
Pricing ShiftGemini 3.7 Flash(Google)Aug 17, 2026
-0.188 $/1M
What Changed

Gemini 3.7 Flash token pricing was updated (-0.188 $/1M).

Why It Matters

Shifts the cost-efficiency curve, enabling higher batch volumes for reduced budget.

Who Benefits

High-volume API users, batch data processors, and agentic loop runners.

Recommendation: Evaluate impact on your current model stack.
Calculate Savings vs Your Model
Pricing ShiftDeepSeek V4 Pro(DeepSeek)Aug 17, 2026
+0.885 $/1M
What Changed

DeepSeek V4 Pro token pricing was updated (+0.885 $/1M).

Why It Matters

Shifts the cost-efficiency curve, enabling higher batch volumes for reduced budget.

Who Benefits

High-volume API users, batch data processors, and agentic loop runners.

Recommendation: Evaluate impact on your current model stack.
Calculate Savings vs Your Model
Pricing ShiftDeepSeek V4 Flash Latest(~deepseek)Aug 17, 2026
-0.007 $/1M
What Changed

DeepSeek V4 Flash Latest token pricing was updated (-0.007 $/1M).

Why It Matters

Shifts the cost-efficiency curve, enabling higher batch volumes for reduced budget.

Who Benefits

High-volume API users, batch data processors, and agentic loop runners.

Recommendation: Evaluate impact on your current model stack.
Calculate Savings vs Your Model
Pricing ShiftMoonshotAI Kimi Latest(~moonshotai)Aug 17, 2026
-1 $/1M
What Changed

MoonshotAI Kimi Latest token pricing was updated (-1 $/1M).

Why It Matters

Shifts the cost-efficiency curve, enabling higher batch volumes for reduced budget.

Who Benefits

High-volume API users, batch data processors, and agentic loop runners.

Recommendation: Evaluate impact on your current model stack.
Calculate Savings vs Your Model
Pricing ShiftCohere: Command R (08-2024)(Cohere)Aug 17, 2026
+9.4 $/1M
What Changed

Cohere: Command R (08-2024) token pricing was updated (+9.4 $/1M).

Why It Matters

Shifts the cost-efficiency curve, enabling higher batch volumes for reduced budget.

Who Benefits

High-volume API users, batch data processors, and agentic loop runners.

Recommendation: Evaluate impact on your current model stack.
Calculate Savings vs Your Model
Pricing ShiftDeepSeek V4 Flash Latest(~deepseek)Aug 17, 2026
-0.014 $/1M
What Changed

DeepSeek V4 Flash Latest token pricing was updated (-0.014 $/1M).

Why It Matters

Shifts the cost-efficiency curve, enabling higher batch volumes for reduced budget.

Who Benefits

High-volume API users, batch data processors, and agentic loop runners.

Recommendation: Evaluate impact on your current model stack.
Calculate Savings vs Your Model
Score MovementQwen2.5 Coder 32B Instruct(Alibaba)Aug 17, 2026
What Changed

Qwen2.5 Coder 32B Instruct demonstrated a score adjustment on evaluations.

Why It Matters

Empirical verification alters the relative head-to-head win-rate against nearest rivals.

Who Benefits

Teams relying on rigorous benchmark boundaries for automated task routing.

Recommendation: Evaluate impact on your current model stack.
Open Head-to-Head Factsheet
Pricing ShiftClaude Opus 5(Anthropic)Aug 17, 2026
-25 $/1M
What Changed

Claude Opus 5 token pricing was updated (-25 $/1M).

Why It Matters

Shifts the cost-efficiency curve, enabling higher batch volumes for reduced budget.

Who Benefits

High-volume API users, batch data processors, and agentic loop runners.

Recommendation: Evaluate impact on your current model stack.
Calculate Savings vs Your Model
Score MovementDeepSeek V4 Flash(DeepSeek)Aug 17, 2026
+262,144 tokens
What Changed

DeepSeek V4 Flash demonstrated a score adjustment on context_window (+262,144 tokens).

Why It Matters

Empirical verification alters the relative head-to-head win-rate against nearest rivals.

Who Benefits

Teams relying on rigorous benchmark boundaries for automated task routing.

Recommendation: Evaluate impact on your current model stack.
Open Head-to-Head Factsheet
Pricing ShiftNVIDIA: Nemotron 3.5 Lightning(nvidia)Aug 17, 2026
-0.05 $/1M
What Changed

NVIDIA: Nemotron 3.5 Lightning token pricing was updated (-0.05 $/1M).

Why It Matters

Shifts the cost-efficiency curve, enabling higher batch volumes for reduced budget.

Who Benefits

High-volume API users, batch data processors, and agentic loop runners.

Recommendation: Evaluate impact on your current model stack.
Calculate Savings vs Your Model
Pricing ShiftQwen: Qwen2.5 VL 72B Instruct(Alibaba)Aug 17, 2026
+0.25 $/1M
What Changed

Qwen: Qwen2.5 VL 72B Instruct token pricing was updated (+0.25 $/1M).

Why It Matters

Shifts the cost-efficiency curve, enabling higher batch volumes for reduced budget.

Who Benefits

High-volume API users, batch data processors, and agentic loop runners.

Recommendation: Evaluate impact on your current model stack.
Calculate Savings vs Your Model
Pricing ShiftGLM-5.2(Zhipu)Aug 17, 2026
+0.952 $/1M
What Changed

GLM-5.2 token pricing was updated (+0.952 $/1M).

Why It Matters

Shifts the cost-efficiency curve, enabling higher batch volumes for reduced budget.

Who Benefits

High-volume API users, batch data processors, and agentic loop runners.

Recommendation: Evaluate impact on your current model stack.
Calculate Savings vs Your Model
Pricing ShiftGLM-5.2(Zhipu)Aug 17, 2026
+0.292 $/1M
What Changed

GLM-5.2 token pricing was updated (+0.292 $/1M).

Why It Matters

Shifts the cost-efficiency curve, enabling higher batch volumes for reduced budget.

Who Benefits

High-volume API users, batch data processors, and agentic loop runners.

Recommendation: Evaluate impact on your current model stack.
Calculate Savings vs Your Model
Pricing ShiftDeepSeek V4 Pro(DeepSeek)Aug 17, 2026
+3.09 $/1M
What Changed

DeepSeek V4 Pro token pricing was updated (+3.09 $/1M).

Why It Matters

Shifts the cost-efficiency curve, enabling higher batch volumes for reduced budget.

Who Benefits

High-volume API users, batch data processors, and agentic loop runners.

Recommendation: Evaluate impact on your current model stack.
Calculate Savings vs Your Model
Score MovementGLM-5.2(Zhipu)Aug 17, 2026
-920,576 tokens
What Changed

GLM-5.2 demonstrated a score adjustment on context_window (-920,576 tokens).

Why It Matters

Empirical verification alters the relative head-to-head win-rate against nearest rivals.

Who Benefits

Teams relying on rigorous benchmark boundaries for automated task routing.

Recommendation: Evaluate impact on your current model stack.
Open Head-to-Head Factsheet
Pricing ShiftGLM-5.2(Zhipu)Aug 17, 2026
+0.882 $/1M
What Changed

GLM-5.2 token pricing was updated (+0.882 $/1M).

Why It Matters

Shifts the cost-efficiency curve, enabling higher batch volumes for reduced budget.

Who Benefits

High-volume API users, batch data processors, and agentic loop runners.

Recommendation: Evaluate impact on your current model stack.
Calculate Savings vs Your Model
Pricing ShiftQwen: Qwen3.6 35B A3B(Alibaba)Aug 17, 2026
-0.01 $/1M
What Changed

Qwen: Qwen3.6 35B A3B token pricing was updated (-0.01 $/1M).

Why It Matters

Shifts the cost-efficiency curve, enabling higher batch volumes for reduced budget.

Who Benefits

High-volume API users, batch data processors, and agentic loop runners.

Recommendation: Evaluate impact on your current model stack.
Calculate Savings vs Your Model
Pricing ShiftDeepSeek V4 Flash(DeepSeek)Aug 17, 2026
-0.057 $/1M
What Changed

DeepSeek V4 Flash token pricing was updated (-0.057 $/1M).

Why It Matters

Shifts the cost-efficiency curve, enabling higher batch volumes for reduced budget.

Who Benefits

High-volume API users, batch data processors, and agentic loop runners.

Recommendation: Evaluate impact on your current model stack.
Calculate Savings vs Your Model
Pricing ShiftQwen: Qwen3 VL 235B A22B Instruct(Alibaba)Aug 17, 2026
-0.05 $/1M
What Changed

Qwen: Qwen3 VL 235B A22B Instruct token pricing was updated (-0.05 $/1M).

Why It Matters

Shifts the cost-efficiency curve, enabling higher batch volumes for reduced budget.

Who Benefits

High-volume API users, batch data processors, and agentic loop runners.

Recommendation: Evaluate impact on your current model stack.
Calculate Savings vs Your Model
Pricing ShiftNVIDIA: Nemotron 3.5 Lightning(nvidia)Aug 17, 2026
-0.02 $/1M
What Changed

NVIDIA: Nemotron 3.5 Lightning token pricing was updated (-0.02 $/1M).

Why It Matters

Shifts the cost-efficiency curve, enabling higher batch volumes for reduced budget.

Who Benefits

High-volume API users, batch data processors, and agentic loop runners.

Recommendation: Evaluate impact on your current model stack.
Calculate Savings vs Your Model