CompareLLM
CompareLLM.ai
Live
Leaderboard
Models
Compare
Best of & Stacks
Research & News
…
CompareLLM
CompareLLM.ai
Precision Benchmarks

Programmatic, dated AI model benchmarks, head-to-head comparisons, and Stack Engine presets.

Daily ingest · 06:00 UTC

Analytics & Benchmarks

  • AI Model Leaderboard
  • Head-to-Head Compare Hub
  • Models Directory
  • Stack Engine Presets
  • Frontier Models
  • Open Weights Catalog

Guides & Intent Lists

  • Best LLM Lists (2026)
  • Best Coding LLM
  • Best Cheap LLM
  • Fastest Low-Latency LLM
  • Claude vs GPT Benchmark
  • What is Elo?
  • Methodology Guides
  • News & Dispatches

Transparency & API

  • Evaluation Methodology
  • Benchmark Changelog
  • Public JSON API
  • llms.txt Specification
  • Privacy Policy
  • Sign In / Account

© 2026 CompareLLM. Public benchmark data aggregated from Arena Elo, LiveBench, SWE-bench & OpenRouter.

Every score has a dated snapshot.

Theme:
Currency:
  1. Home
  2. Stack Engine
  3. long-context RAG
Deterministic Stack Engine

Best AI Stack for long-context RAG

Context window and input price for stuffing large corpora.

Objective Weight Distribution
Context:40%In $:25%Elo:20%TTFT:15%
Recommended PickScore 82.8 / 100

GPT-5.6 Luna

OpenAI · Closed flagship

Context: 1.1M (Top 5%)In $: $0.1/1M (Top 19%)Elo: 1,466 (Top 51%)TTFT: 88 ms (Top 2%)
Output: $0.6/1MSWE-bench: 61.8%
View full model fact sheet

Full Stack Leaderboard

Ranked alternatives optimized for long-context RAG based on objective benchmark weighting.

#2DeepSeek V4 Flash(DeepSeek)
Score 81.6
Context1.3MIn $$0.14/1MElo1,490TTFT165 ms
Rank #2 in stackvs #1
#3Llama 4 Scout(Meta)
Score 78.6
Context1.3MIn $$0.1/1MElo1,412TTFT130 ms
Rank #3 in stackvs #1
#4Gemini 3.7 Flash(Google)
Score 78.3
Context1MIn $$0.375/1MElo1,530TTFT82 ms
Rank #4 in stackvs #1
#5Gemini 2.0 Flash Thinking(Google)
Score 74.6
Context1MIn $$0.1/1MElo1,490TTFT220 ms
Rank #5 in stackvs #1
#6Gemini 3 Flash(Google)
Score 74.3
Context1MIn $$0.5/1MElo1,495TTFT95 ms
Rank #6 in stackvs #1
#7GPT-5.6 Terra(OpenAI)
Score 73.8
Context1.1MIn $$1/1MElo1,548TTFT160 ms
Rank #7 in stackvs #1
#8Llama 4 Maverick(Meta)
Score 73.0
Context1MIn $$0.2/1MElo1,468TTFT170 ms
Rank #8 in stackvs #1
#9Gemini 3.6 Pro(Google)
Score 72.8
Context2MIn $$1.25/1MElo1,570TTFT205 ms
Rank #9 in stackvs #1
#10Gemini 3.6 Flash(Google)
Score 72.8
Context1MIn $$0.75/1MElo1,506TTFT92 ms
Rank #10 in stackvs #1