CompareLLM
CompareLLM.ai
Live
LeaderboardModelsCompareStacksBest ofGuidesNewsMethod
…
CompareLLM
CompareLLM.ai
Precision Benchmarks

Programmatic, dated AI model benchmarks, head-to-head comparisons, and Stack Engine presets.

Daily ingest · 06:00 UTC

Analytics & Benchmarks

  • AI Model Leaderboard
  • Head-to-Head Compare Hub
  • Models Directory
  • Stack Engine Presets
  • Frontier Models
  • Open Weights Catalog

Guides & Intent Lists

  • Best LLM Lists (2026)
  • Best Coding LLM
  • Best Cheap LLM
  • Fastest Low-Latency LLM
  • Claude vs GPT Benchmark
  • What is Elo?
  • Methodology Guides
  • News & Dispatches

Transparency & API

  • Evaluation Methodology
  • Benchmark Changelog
  • Public JSON API
  • llms.txt Specification
  • Privacy Policy
  • Sign In / Account

© 2026 CompareLLM. Public benchmark data aggregated from Arena Elo, LiveBench, SWE-bench & OpenRouter.

Every score has a dated snapshot.

  1. Home
  2. Providers
  3. DeepSeek
AI Provider Portfolio
12 active models in catalog

DeepSeek Models

Comprehensive benchmark scores, coding performance, latency metrics, and API pricing for all models developed by DeepSeek.

Flagship ModelDeepSeek V4 Pro
Portfolio Avg Elo1412 Elo
Browse Other Providers
AnthropicOpenAIGoogle

Available Models & Verified Benchmarks

#1DeepSeek V4 Pro
Open Weights

Open-weight-adjacent DeepSeek flagship. High reasoning density per dollar.

Elo1,536
SWE-bench71.6%
Out $/1M$0.87
Model Fact Sheetvs all hub
#2DeepSeek V4 Flash
Open Weights

DeepSeek Jul 31 2026 price-performance SKU. Public listings put it near $0.14/$0.28 per 1M tokens.

Elo1,490
SWE-bench71.2%
Out $/1M$0.28
Model Fact Sheetvs all hub
#3DeepSeek Coder V2
Open Weights

DeepSeek open-weight Mixture-of-Experts coding model supporting 338 programming languages and 128k context.

Elo1,365
SWE-bench60.5%
Out $/1M$0.28
Model Fact Sheetvs all hub
#4DeepSeek R1
Open Weights

Open-weights reasoning model trained with large-scale RL.

Elo1,358
SWE-bench65.2%
Out $/1M$2.5
Model Fact Sheetvs all hub
#5DeepSeek V3
Open Weights

Prior DeepSeek flagship. Baseline for v3 vs v4.

Elo1,310
SWE-bench48.6%
Out $/1M$1.0287
Model Fact Sheetvs all hub
#6DeepSeek: DeepSeek V3.2
Proprietary

Auto-discovered from OpenRouter (deepseek/deepseek-v3.2). Preview until a second source matches.

Elo—
SWE-bench—
Out $/1M$0.39999999999999997
Model Fact Sheetvs all hub
#7DeepSeek: DeepSeek V3.2 Exp
Proprietary

Auto-discovered from OpenRouter (deepseek/deepseek-v3.2-exp). Preview until a second source matches.

Elo—
SWE-bench—
Out $/1M$0.41
Model Fact Sheetvs all hub
#8DeepSeek: DeepSeek V3.1 Terminus
Proprietary

Auto-discovered from OpenRouter (deepseek/deepseek-v3.1-terminus). Preview until a second source matches.

Elo—
SWE-bench—
Out $/1M$0.95
Model Fact Sheetvs all hub
#9DeepSeek: DeepSeek V3.1
Proprietary

Auto-discovered from OpenRouter (deepseek/deepseek-chat-v3.1). Preview until a second source matches.

Elo—
SWE-bench—
Out $/1M$0.95
Model Fact Sheetvs all hub
#10DeepSeek: R1 0528
Proprietary

Auto-discovered from OpenRouter (deepseek/deepseek-r1-0528). Preview until a second source matches.

Elo—
SWE-bench—
Out $/1M$2.1500000000000004
Model Fact Sheetvs all hub
#11DeepSeek: DeepSeek V3 0324
Proprietary

Auto-discovered from OpenRouter (deepseek/deepseek-chat-v3-0324). Preview until a second source matches.

Elo—
SWE-bench—
Out $/1M$1.12
Model Fact Sheetvs all hub
#12DeepSeek: R1 Distill Llama 70B
Proprietary

Auto-discovered from OpenRouter (deepseek/deepseek-r1-distill-llama-70b). Preview until a second source matches.

Elo—
SWE-bench—
Out $/1M$0.7999999999999999
Model Fact Sheetvs all hub