Frontier AI Models
30 state-of-the-art models actively pushing the frontier in reasoning, coding, and autonomous benchmark scores.
Anthropic current default flagship. Official API $5/$25 per 1M tokens and a 1M context window (Anthropic, Jul 24 2026).
Anthropic top-tier long-horizon model. Official API $10/$50 per 1M tokens (Claude Platform pricing, Aug 2026).
OpenAI 5.6 flagship tier. Official API $5/$30 per 1M tokens (OpenAI pricing, Jul 30 2026 update left Sol unchanged).
Prior Opus generation still billed at $5/$25. Kept as a compare baseline against Opus 5.
xAI Aug 12 2026 post-training refresh of Grok 4.5. Same $2/$6 API price, 500k context, stronger agentic traces.
Follow-on Opus release. Slightly behind 4.5 on official SWE-bench bash-only in the last published sweep; stronger long-horizon agent traces.
Current Google Pro-class multimodal model. Long context, strong coding, billed like the 3.x Pro tier.
Anthropic frontier coding and computer-use model. SWE-bench leader on the official mini-SWE-agent harness in the Feb 2026 refresh.
OpenAI flagship reasoning model for 2025–26. Strong general preference Elo and multimodal coverage.
Z.ai Aug 14 2026 post-train of the GLM-5.2 744B base. Coding-plan live; open weights promised after a two-week safety review (z.ai/blog/glm-5.3).
Anthropic workhorse. Official $2/$10 per 1M tokens made permanent on Aug 10 2026.
Google frontier multimodal model with a multi-million-token context window.
OpenAI 5.6 mid tier. Official API $2/$12 per 1M after the Jul 30 2026 price cut.
OpenAI largest dense non-reasoning model with expansive world knowledge and reduced hallucinations.
Open-weight-adjacent DeepSeek flagship. High reasoning density per dollar.
Google Aug 13 2026 workhorse. Official intro price $0.75/$3.75 per 1M through Dec 31 2026; 1,048,576-token context (Google blog).
Moonshot AI frontier flagship reasoning and agent swarm model with 256k context and top-tier SWE-bench coding capability.
Workhorse Anthropic model: most of Opus coding quality at a mid-tier price.
Alibaba flagship. Strong math and multilingual code.
Google Jul 21 workhorse. Now shares the 3.7 Flash introductory $0.75/$3.75 rate through Dec 31 2026.
Fast Google frontier-adjacent model. Near-top official SWE-bench at a fraction of Opus price.
Alibaba specialized open reasoning model competing with frontier closed reasoning models.
Zhipu flagship. Strong Chinese/English coding and agents.
Google experimental reasoning model that visualizes thoughts in real-time.
DeepSeek Jul 31 2026 price-performance SKU. Public listings put it near $0.14/$0.28 per 1M tokens.
MiniMax coding model. Tied near the top of official SWE-bench bash-only in Feb 2026.
01.AI ultra-fast reasoning model delivering top LiveBench efficiency.
ByteDance flagship enterprise model with ultra-low token cost and 128k context.
Meta natively multimodal open-weight flagship.
Mistral European flagship with strong function calling.
