CompareLLM
CompareLLM.ai
Live
LeaderboardModelsCompareStacksBest ofGuidesNewsMethod
…
CompareLLM
CompareLLM.ai
Precision Benchmarks

Programmatic, dated AI model benchmarks, head-to-head comparisons, and Stack Engine presets.

Daily ingest · 06:00 UTC

Analytics & Benchmarks

  • AI Model Leaderboard
  • Head-to-Head Compare Hub
  • Models Directory
  • Stack Engine Presets
  • Frontier Models
  • Open Weights Catalog

Guides & Intent Lists

  • Best LLM Lists (2026)
  • Best Coding LLM
  • Best Cheap LLM
  • Fastest Low-Latency LLM
  • Claude vs GPT Benchmark
  • What is Elo?
  • Methodology Guides
  • News & Dispatches

Transparency & API

  • Evaluation Methodology
  • Benchmark Changelog
  • Public JSON API
  • llms.txt Specification
  • Privacy Policy
  • Sign In / Account

© 2026 CompareLLM. Public benchmark data aggregated from Arena Elo, LiveBench, SWE-bench & OpenRouter.

Every score has a dated snapshot.

  1. Home
  2. Best lists
  3. best llm for rag
Search Intent · best llm for rag

Best LLM for RAG in 2026

Context window and input list price first. For stuffing corpora, not for coding agents.

Quick answer

GPT-5.6 Luna is the current #1 for “best llm for rag” on this dated Stack Engine mix. Context window and input price for stuffing large corpora. Weights: Context 40%, In $ 25%, Elo 20%, TTFT 15%.
Weights:Context 40%In $ 25%Elo 20%TTFT 15%
Current #1 Ranked PickScore 82.8 / 100

GPT-5.6 Luna

OpenAI · Closed flagship

context window p95input price per m p81elo p49ttft ms p98
SWE-bench: 61.8%Elo: 1,466
View full model fact sheet

Complete Ranked Category List

Models ranked by verified benchmark weights across SWE-bench and coding preference evaluations.

#1GPT-5.6 Luna(OpenAI)
Score 82.8
context windowP95input price per mP81eloP49ttft msP98
🏆 #1 Overall Leader in Category
#2Llama 4 Scout(Meta)
Score 78.8
context windowP98input price per mP81eloP32ttft msP86
Rank #2vs #1
#3Gemini 3.7 Flash(Google)
Score 78.1
context windowP87input price per mP53eloP76ttft msP99
Rank #3vs #1
#4DeepSeek V4 Flash(DeepSeek)
Score 77.0
context windowP87input price per mP75eloP61ttft msP75
Rank #4vs #1
#5Gemini 2.0 Flash Thinking(Google)
Score 74.4
context windowP87input price per mP79eloP61ttft msP51
Rank #5vs #1
#6Gemini 3 Flash(Google)
Score 73.9
context windowP87input price per mP47eloP66ttft msP94
Rank #6vs #1
#7GPT-5.6 Terra(OpenAI)
Score 73.8
context windowP95input price per mP32eloP81ttft msP77
Rank #7vs #1
#8Gemini 3.6 Pro(Google)
Score 73.0
context windowP99input price per mP26eloP91ttft msP58
Rank #8vs #1
#9Llama 4 Maverick(Meta)
Score 73.0
context windowP87input price per mP68eloP51ttft msP73
Rank #9vs #1
#10Gemini 3.6 Flash(Google)
Score 72.3
context windowP87input price per mP38eloP68ttft msP96
Rank #10vs #1
#11Kimi Chat 1.5(Moonshot)
Score 71.6
context windowP99input price per mP67eloP34ttft msP56
Rank #11vs #1
#12Gemini 3 Pro(Google)
Score 70.2
context windowP99input price per mP26eloP82ttft msP51
Rank #12vs #1

Frequently asked questions

Plain-English methodology and leaderboard answers

GPT-5.6 Luna is the current #1 on this list. Rankings move when daily ingest updates SWE-bench, Elo, price, or latency.

What is preference Elo? · How rankings update