CompareLLM
CompareLLM.ai
Live
Leaderboard
Models
Compare
Best of & Stacks
Research & News
…
CompareLLM
CompareLLM.ai
Precision Benchmarks

Programmatic, dated AI model benchmarks, head-to-head comparisons, and Stack Engine presets.

Daily ingest · 06:00 UTC

Analytics & Benchmarks

  • AI Model Leaderboard
  • Head-to-Head Compare Hub
  • Models Directory
  • Stack Engine Presets
  • Frontier Models
  • Open Weights Catalog

Guides & Intent Lists

  • Best LLM Lists (2026)
  • Best Coding LLM
  • Best Cheap LLM
  • Fastest Low-Latency LLM
  • Claude vs GPT Benchmark
  • What is Elo?
  • Methodology Guides
  • News & Dispatches

Transparency & API

  • Evaluation Methodology
  • Benchmark Changelog
  • Public JSON API
  • llms.txt Specification
  • Privacy Policy
  • Sign In / Account

© 2026 CompareLLM. Public benchmark data aggregated from Arena Elo, LiveBench, SWE-bench & OpenRouter.

Every score has a dated snapshot.

Theme:
Currency:
  1. Home
  2. News
  3. DeepSeek V4 Flash (Jul 31): the $0.14/$0.28 price-performance SKU
DeepSeek V4 Flash (Jul 31): the $0.14/$0.28 price-performance SKU
priceVerified Dispatch
CompareLLM Intelligence Desk·Jul 31, 2026

DeepSeek V4 Flash (Jul 31): the $0.14/$0.28 price-performance SKU

Public listings put Flash near $0.14/$0.28 per 1M. That is a batch SKU, not a chatbot crown.

Search intent: deepseek v4 flash price

Evaluated Models (2):
DeepSeek V4 FlashDeepSeek
DeepSeek V4 ProDeepSeek
Share Analysis:
WhatsAppTelegramXLinkedInReddit
Verified Benchmark Scorecard2 Models Evaluated

DeepSeek V4 Flash vs DeepSeek V4 Pro: Live Benchmark Matchup

Live benchmark scores, throughput speeds, and token pricing with dynamic peer comparison.

Full Head-to-Head
LMSYS Arena Elo1,490DeepSeek V4 Flash
SWE-bench Verified71.2%Coding resolve %
Throughput Speed142 tok/sGeneration rate
Output Price / 1M$0.28List API rate

93% Output Token Cost Advantage

DeepSeek V4 Flash costs $0.28 / 1M tokens compared to DeepSeek V4 Pro at $3.96 / 1M tokens ($3.68/1M tokens difference).

Interactive Model Comparison:2 models selected
DeepSeek V4 FlashArticleDeepSeek V4 ProArticle
LMSYS Arena EloOverall human preference
quality
DeepSeek V4 Flash
1,490
DeepSeek V4 ProTop
1,536
Coding EloProgramming preference
quality
DeepSeek V4 Flash
1,516
DeepSeek V4 ProTop
1,544
SWE-bench VerifiedGitHub issue resolve %
quality
DeepSeek V4 Flash
71.2%
DeepSeek V4 ProTop
71.6%
GPQA DiamondPhD-level science reasoning %
quality
DeepSeek V4 Flash
81.6%
DeepSeek V4 ProTop
84%
LiveBenchContamination-free reasoning
quality
DeepSeek V4 Flash
67.8%
DeepSeek V4 ProTop
70.1%
Throughput (Speed)Output tokens / second
speed
DeepSeek V4 FlashTop
142 tok/s
DeepSeek V4 Pro
70 tok/s
Time-to-First-TokenInitial latency (ms)
speed
DeepSeek V4 FlashTop
165 ms
DeepSeek V4 Pro
410 ms
Output Token Price$ per 1M reply tokens
price
DeepSeek V4 FlashTop
$0.28/1M
DeepSeek V4 Pro
$3.96/1M
Input Token Price$ per 1M prompt tokens
price
DeepSeek V4 FlashTop
$0.14/1M
DeepSeek V4 Pro
$1.32/1M
Context WindowMax token capacity
capacity
DeepSeek V4 FlashTop
1.3M
DeepSeek V4 Pro
1M
Benchmark / Metric
DeepSeek V4 FlashDeepSeek · Article Model
DeepSeek V4 ProDeepSeek · Article Model
LMSYS Arena EloOverall human preference
1,490
1,536Top
Coding EloProgramming preference
1,516
1,544Top
SWE-bench VerifiedGitHub issue resolve %
71.2%
71.6%Top
GPQA DiamondPhD-level science reasoning %
81.6%
84%Top
LiveBenchContamination-free reasoning
67.8%
70.1%Top
Throughput (Speed)Output tokens / second
142 tok/sTop
70 tok/s
Time-to-First-TokenInitial latency (ms)
165 msTop
410 ms
Output Token Price$ per 1M reply tokens
$0.28/1MTop
$3.96/1M
Input Token Price$ per 1M prompt tokens
$0.14/1MTop
$1.32/1M
Context WindowMax token capacity
1.3MTop
1M
Executive Key Takeaway
Public listings put Flash near $0.14/$0.28 per 1M. That is a batch SKU, not a chatbot crown.

DeepSeek V4 Flash is the Jul 31 2026 price-performance SKU. Public listings put it near $0.14/$0.28 per million tokens. On CompareLLM that is a cheapest-api and cheap-coding candidate, not a general-assistant trophy. Compare it to Luna and Gemini Flash on list price and SWE-bench, then measure retries on your own traffic.

Ultra-cheap output still loses to retries

If Flash needs three samples to match Sonnet on your tests, the invoice flips. Pair-page sketches assume one successful call.

Empirical Evaluation & Architectural Analysis

Empirical evaluations recorded across the CompareLLM Leaderboard highlight how parameter scaling and inference optimization impact production throughput and cost-per-token economics. Reference the interactive scorecard above for verified snapshots or inspect the dedicated deepseek-v4-flash dossier.

Keep Flash and Pro as two slugs

V4 Pro and V4 Flash must not share a canonical id. Aliases like “v4-flash” stay on Flash only.

Strategic Deployment Recommendation

  • Production Workloads: For enterprise workloads requiring strict reliability, cross-reference Top Coding LLMs and Cheapest High-Quality LLMs.
  • Direct Pair Comparison: Explore the live pairwise breakdown at /compare/deepseek-v4-flash-vs-gpt-5-6-luna.
  • Architectural Stacks: Recommended architectural configurations can be evaluated on the CompareLLM Stack Engine.

Frequently Asked Questions

Is this an CompareLLM lab test score? No. Linked model dossiers and compare showdowns reflect dated snapshots gathered from named public evaluation benchmarks. Seed catalog entries remain transparently timestamped until daily ingest updates them.

Where can I see live comparisons for this model? View the showdown at /compare/deepseek-v4-flash-vs-gpt-5-6-luna.

What primary search query does this briefing answer? deepseek v4 flash price.

Share Analysis:
WhatsAppTelegramXLinkedInReddit

Head-to-Head Showdowns for Mentioned Models

Featured Article ShowdownDeepSeek V4 Flash vs DeepSeek V4 Pro
Compare Benchmarks
Benchmark MatchupDeepSeek V4 Flash vs Claude Opus 4.5
Benchmark MatchupDeepSeek V4 Flash vs Claude Opus 4.6
Benchmark MatchupDeepSeek V4 Pro vs Claude Opus 4.5
Benchmark MatchupDeepSeek V4 Pro vs Claude Opus 4.6

Related news

  • compare · Jun 13, 2026

    DeepSeek vs Claude: the cheap-versus-frontier question, remapped

  • launch · Jun 12, 2026

    DeepSeek V4 Pro: high reasoning density per dollar

  • launch · Jan 20, 2025

    DeepSeek R1 stays as the open-weights RL baseline

Back to All DispatchesExplore Comparison Matrix