CompareLLM
CompareLLM.ai
Live
LeaderboardModelsCompareStacksBest ofGuidesNewsMethod
…
CompareLLM
CompareLLM.ai
Precision Benchmarks

Programmatic, dated AI model benchmarks, head-to-head comparisons, and Stack Engine presets.

Daily ingest · 06:00 UTC

Analytics & Benchmarks

  • AI Model Leaderboard
  • Head-to-Head Compare Hub
  • Models Directory
  • Stack Engine Presets
  • Frontier Models
  • Open Weights Catalog

Guides & Intent Lists

  • Best LLM Lists (2026)
  • Best Coding LLM
  • Best Cheap LLM
  • Fastest Low-Latency LLM
  • Claude vs GPT Benchmark
  • What is Elo?
  • Methodology Guides
  • News & Dispatches

Transparency & API

  • Evaluation Methodology
  • Benchmark Changelog
  • Public JSON API
  • llms.txt Specification
  • Privacy Policy
  • Sign In / Account

© 2026 CompareLLM. Public benchmark data aggregated from Arena Elo, LiveBench, SWE-bench & OpenRouter.

Every score has a dated snapshot.

  1. Home
  2. Providers
  3. xAI
AI Provider Portfolio
7 active models in catalog

xAI Models

Comprehensive benchmark scores, coding performance, latency metrics, and API pricing for all models developed by xAI.

Flagship ModelGrok 4.6
Portfolio Avg Elo1495 Elo
Browse Other Providers
AnthropicOpenAIGoogle

Available Models & Verified Benchmarks

#1Grok 4.6
Proprietary

xAI Aug 12 2026 post-training refresh of Grok 4.5. Same $2/$6 API price, 500k context, stronger agentic traces.

Elo1,592
SWE-bench69.1%
Out $/1M$6
Model Fact Sheetvs all hub
#2Grok 4
Proprietary

Previous xAI flagship.

Elo1,488
SWE-bench58.6%
Out $/1M$15
Model Fact Sheetvs all hub
#3Grok 3
Proprietary

Previous xAI generation before Grok 4. Kept for grok 3 vs grok 4.

Elo1,405
SWE-bench51.4%
Out $/1M$15
Model Fact Sheetvs all hub
#4SpaceXAI: Grok Build 0.1
Proprietary

Auto-discovered from OpenRouter (x-ai/grok-build-0.1). Preview until a second source matches.

Elo—
SWE-bench—
Out $/1M$2
Model Fact Sheetvs all hub
#5SpaceXAI: Grok 4.3
Proprietary

Auto-discovered from OpenRouter (x-ai/grok-4.3). Preview until a second source matches.

Elo—
SWE-bench—
Out $/1M$2.5
Model Fact Sheetvs all hub
#6SpaceXAI: Grok 4.20 Multi-Agent
Proprietary

Auto-discovered from OpenRouter (x-ai/grok-4.20-multi-agent). Preview until a second source matches.

Elo—
SWE-bench—
Out $/1M$2.5
Model Fact Sheetvs all hub
#7SpaceXAI: Grok 4.20
Proprietary

Auto-discovered from OpenRouter (x-ai/grok-4.20). Preview until a second source matches.

Elo—
SWE-bench—
Out $/1M$2.5
Model Fact Sheetvs all hub