CompareLLM
CompareLLM.ai
Live
LeaderboardModelsCompareStacksBest ofGuidesNewsMethod
…
CompareLLM
CompareLLM.ai
Precision Benchmarks

Programmatic, dated AI model benchmarks, head-to-head comparisons, and Stack Engine presets.

Daily ingest · 06:00 UTC

Analytics & Benchmarks

  • AI Model Leaderboard
  • Head-to-Head Compare Hub
  • Models Directory
  • Stack Engine Presets
  • Frontier Models
  • Open Weights Catalog

Guides & Intent Lists

  • Best LLM Lists (2026)
  • Best Coding LLM
  • Best Cheap LLM
  • Fastest Low-Latency LLM
  • Claude vs GPT Benchmark
  • What is Elo?
  • Methodology Guides
  • News & Dispatches

Transparency & API

  • Evaluation Methodology
  • Benchmark Changelog
  • Public JSON API
  • llms.txt Specification
  • Privacy Policy
  • Sign In / Account

© 2026 CompareLLM. Public benchmark data aggregated from Arena Elo, LiveBench, SWE-bench & OpenRouter.

Every score has a dated snapshot.

  1. Home
  2. Models
  3. Gemini 3.7 Flash
GoogleFrontier Model

Gemini 3.7 Flash

Gemini 3.7 Flash is a Google closed-API frontier model. Latest preference Elo in this catalog is 1,530 (seed-bootstrap (Aug 16, 2026)). SWE-bench sits at 72.4% (seed-bootstrap (Aug 16, 2026)). List output price is $1.88/1M. Context window is 1M tokens. Google Aug 13 2026 workhorse. Official intro price $0.75/$3.75 per 1M through Dec 31 2026; 1,048,576-token context (Google blog). Numbers below are dated snapshots, not a guarantee on your traffic mix.

At a glance

Gemini 3.7 Flash is a Google closed-API frontier model. Latest preference Elo in this catalog is 1,530 (seed-bootstrap (Aug 16, 2026)). SWE-bench sits at 72.4% (seed-bootstrap (Aug 16, 2026)). List output price is $1.88/1M. Context window is 1M tokens. Google Aug 13 2026 workhorse. Official intro price $0.75/$3.75 per 1M through Dec 31 2026; 1,048,576-token context (Google blog). Numbers below are dated snapshots, not a guarantee on your traffic mix.
Last updated Aug 16, 2026
All 69 vsSave Model

Compare Gemini 3.7 Flash against

Capability Profile

Gemini 3.7 Flash Benchmark Percentiles

Plotted against all active catalog models (50th percentile = catalog median).

Gemini 3.7 Flash
Median (50)

Dimensional Scorecards

Reasoning
75%
Coding
77%
Math
74%
Speed
99%
Value
49%
Context
87%
Standout Competencies

Ranks in the top tier (≥75th percentile) for Speed, Context, Coding, Reasoning.

Benchmark Specifications

Dated snapshot metrics aggregated from official evaluators and API providers.

Preference Elo1,530
seed-bootstrap · Aug 16, 2026
Coding Elo1,546
seed-bootstrap · Aug 16, 2026
LiveBench69.6%
seed-bootstrap · Aug 16, 2026
SWE-bench72.4%
seed-bootstrap · Aug 16, 2026
GPQA Diamond81.2%
seed-bootstrap · Aug 16, 2026
Time to first token82 ms
seed-bootstrap · Aug 16, 2026
Output speed215 tok/s
seed-bootstrap · Aug 16, 2026
Input price$0.38/1M
openrouter · Aug 16, 2026
Output price$1.88/1M
openrouter · Aug 16, 2026
Context window1M
openrouter · Aug 16, 2026
Benchmark MetricReported ScoreObserved Source
Preference Elo1,530seed-bootstrap · Aug 16, 2026
Coding Elo1,546seed-bootstrap · Aug 16, 2026
LiveBench69.6%seed-bootstrap · Aug 16, 2026
SWE-bench72.4%seed-bootstrap · Aug 16, 2026
GPQA Diamond81.2%seed-bootstrap · Aug 16, 2026
Time to first token82 msseed-bootstrap · Aug 16, 2026
Output speed215 tok/sseed-bootstrap · Aug 16, 2026
Input price$0.38/1Mopenrouter · Aug 16, 2026
Output price$1.88/1Mopenrouter · Aug 16, 2026
Context window1Mopenrouter · Aug 16, 2026

Compare with every model

Search or pick any catalog row. Suggested matchups first, then the full list.

Dedicated vs hub →
Gemini 3.7 Flash vs Gemini 3.6 Flash (previous gemini-flash)Gemini 3.7 Flash vs Kimi K3Gemini 3.7 Flash vs Claude Sonnet 4.5Gemini 3.7 Flash vs DeepSeek V4 ProGemini 3.7 Flash vs GPT-4.5 OrionGemini 3.7 Flash vs Qwen 3 Max
Showing 69 of 69 comparisons
Gemini 3.7 Flash vs Claude Opus 5Anthropic
Compare

Anthropic current default flagship. Official API $5/$25 per 1M tokens and a 1M context window (Anthropic, Jul 24 2026).

Elo: −94SWE: −6.8%
Gemini 3.7 Flash vs Claude Fable 5Anthropic
Compare

Anthropic top-tier long-horizon model. Official API $10/$50 per 1M tokens (Claude Platform pricing, Aug 2026).

Elo: −86SWE: −7.6%
Gemini 3.7 Flash vs GPT-5.6 SolOpenAI
Compare

OpenAI 5.6 flagship tier. Official API $5/$30 per 1M tokens (OpenAI pricing, Jul 30 2026 update left Sol unchanged).

Elo: −78SWE: −5.2%
Gemini 3.7 Flash vs Claude Opus 4.8Anthropic
Compare

Prior Opus generation still billed at $5/$25. Kept as a compare baseline against Opus 5.

Elo: −68SWE: −6%
Gemini 3.7 Flash vs Grok 4.6xAI
Compare

xAI Aug 12 2026 post-training refresh of Grok 4.5. Same $2/$6 API price, 500k context, stronger agentic traces.

Elo: −62SWE: +3.3%
Gemini 3.7 Flash vs Claude Opus 4.6Anthropic
Compare

Follow-on Opus release. Slightly behind 4.5 on official SWE-bench bash-only in the last published sweep; stronger long-horizon agent traces.

Elo: −44SWE: −3.5%
Gemini 3.7 Flash vs Gemini 3.6 ProGoogle
Compare

Current Google Pro-class multimodal model. Long context, strong coding, billed like the 3.x Pro tier.

Elo: −40SWE: −2.2%
Gemini 3.7 Flash vs Claude Opus 4.5Anthropic
Compare

Anthropic frontier coding and computer-use model. SWE-bench leader on the official mini-SWE-agent harness in the Feb 2026 refresh.

Elo: −38SWE: −4.4%
Gemini 3.7 Flash vs OpenAI: o3 MiniOpenAI
Compare

Auto-discovered from OpenRouter (openai/o3-mini). Preview until a second source matches.

Elo: −30SWE: −6.1%
Gemini 3.7 Flash vs GPT-5OpenAI
Compare

OpenAI flagship reasoning model for 2025–26. Strong general preference Elo and multimodal coverage.

Elo: −28SWE: +4%
Gemini 3.7 Flash vs GLM-5.3Zhipu
Compare

Z.ai Aug 14 2026 post-train of the GLM-5.2 744B base. Coding-plan live; open weights promised after a two-week safety review (z.ai/blog/glm-5.3).

Elo: −28SWE: −4%
Gemini 3.7 Flash vs Claude Sonnet 5Anthropic
Compare

Anthropic workhorse. Official $2/$10 per 1M tokens made permanent on Aug 10 2026.

Elo: −26SWE: −1.4%
Gemini 3.7 Flash vs Gemini 3 ProGoogle
Compare

Google frontier multimodal model with a multi-million-token context window.

Elo: −22SWE: +0%
Gemini 3.7 Flash vs GPT-5.6 TerraOpenAI
Compare

OpenAI 5.6 mid tier. Official API $2/$12 per 1M after the Jul 30 2026 price cut.

Elo: −18SWE: +2%
Gemini 3.7 Flash vs GPT-4.5 OrionOpenAI
Compare

OpenAI largest dense non-reasoning model with expansive world knowledge and reduced hallucinations.

Elo: −10SWE: +1.4%
Gemini 3.7 Flash vs DeepSeek V4 ProDeepSeek
Compare

Open-weight-adjacent DeepSeek flagship. High reasoning density per dollar.

Elo: −6SWE: +0.8%
Gemini 3.7 Flash vs Kimi K3Moonshot
Compare

Moonshot AI frontier flagship reasoning and agent swarm model with 256k context and top-tier SWE-bench coding capability.

Elo: +2SWE: −1.4%
Gemini 3.7 Flash vs Claude Sonnet 4.5Anthropic
Compare

Workhorse Anthropic model: most of Opus coding quality at a mid-tier price.

Elo: +6SWE: +2.3%
Gemini 3.7 Flash vs Qwen 3 MaxAlibaba
Compare

Alibaba flagship. Strong math and multilingual code.

Elo: +12SWE: +4.6%
Gemini 3.7 Flash vs MoonshotAI: Kimi K2.5Moonshot
Compare

Auto-discovered from OpenRouter (moonshotai/kimi-k2.5). Preview until a second source matches.

Elo: +15SWE: +1.1%
Gemini 3.7 Flash vs OpenAI: o1OpenAI
Compare

Auto-discovered from OpenRouter (openai/o1). Preview until a second source matches.

Elo: +18SWE: +23.5%
Gemini 3.7 Flash vs Gemini 3.6 FlashGoogle
Compare

Google Jul 21 workhorse. Now shares the 3.7 Flash introductory $0.75/$3.75 rate through Dec 31 2026.

Elo: +24SWE: +1.6%
Gemini 3.7 Flash vs Gemini 3 FlashGoogle
Compare

Fast Google frontier-adjacent model. Near-top official SWE-bench at a fraction of Opus price.

Elo: +35SWE: −3.4%
Gemini 3.7 Flash vs Qwen QwQ 32BAlibaba
Compare

Alibaba specialized open reasoning model competing with frontier closed reasoning models.

Elo: +35SWE: +4.9%
Gemini 3.7 Flash vs GLM-5.2Zhipu
Compare

Zhipu flagship. Strong Chinese/English coding and agents.

Elo: +38SWE: +3.1%
Gemini 3.7 Flash vs Gemini 2.0 Flash ThinkingGoogle
Compare

Google experimental reasoning model that visualizes thoughts in real-time.

Elo: +40SWE: +5.6%
Gemini 3.7 Flash vs DeepSeek V4 FlashDeepSeek
Compare

DeepSeek Jul 31 2026 price-performance SKU. Public listings put it near $0.14/$0.28 per 1M tokens.

Elo: +40SWE: +1.2%
Gemini 3.7 Flash vs Claude Opus 4Anthropic
Compare

First Claude 4 Opus generation. Baseline for opus 4 vs opus 5.

Elo: +40SWE: −0.1%
Gemini 3.7 Flash vs Grok 4xAI
Compare

Previous xAI flagship.

Elo: +42SWE: +13.8%
Gemini 3.7 Flash vs MiniMax M2.5MiniMax
Compare

MiniMax coding model. Tied near the top of official SWE-bench bash-only in Feb 2026.

Elo: +52SWE: −3.4%
Gemini 3.7 Flash vs Yi-Lightning01.AI
Compare

01.AI ultra-fast reasoning model delivering top LiveBench efficiency.

Elo: +55SWE: +9%
Gemini 3.7 Flash vs Seed 2.1 TurboByteDance
Compare

ByteDance Seed 2.1 Turbo, listed on public model timelines as an Aug 10 2026 API drop.

Elo: +58SWE: +9.2%
Gemini 3.7 Flash vs Doubao Pro 1.5ByteDance
Compare

ByteDance flagship enterprise model with ultra-low token cost and 128k context.

Elo: +60SWE: +10.3%
Gemini 3.7 Flash vs Llama 4 MaverickMeta
Compare

Meta natively multimodal open-weight flagship.

Elo: +62SWE: +15.3%
Gemini 3.7 Flash vs GPT-5.6 LunaOpenAI
Compare

OpenAI 5.6 fast/cheap tier. Official API $0.20/$1.20 per 1M after the Jul 30 80% Luna cut.

Elo: +64SWE: +10.6%
Gemini 3.7 Flash vs MoonshotAI: Kimi K2 0711Moonshot
Compare

Auto-discovered from OpenRouter (moonshotai/kimi-k2). Preview until a second source matches.

Elo: +65SWE: +6.6%
Gemini 3.7 Flash vs Mistral Large 3Mistral
Compare

Mistral European flagship with strong function calling.

Elo: +74SWE: +17%
Gemini 3.7 Flash vs Claude Sonnet 4Anthropic
Compare

First Sonnet 4 generation. Bridge between 3.5/3.7 and Sonnet 5.

Elo: +75SWE: +9.6%
Gemini 3.7 Flash vs Ernie 4.5 TurboBaidu
Compare

Baidu current multimodal enterprise foundation model with broad Chinese knowledge.

Elo: +80SWE: +15.4%
Gemini 3.7 Flash vs Qwen 2.5 PlusAlibaba
Compare

Alibaba balanced flagship API model with high-throughput general reasoning.

Elo: +85SWE: +13.2%
Gemini 3.7 Flash vs OpenAI o1-miniOpenAI
Compare

OpenAI high-speed, cost-effective reasoning model optimized for STEM, math, and code generation.

Elo: +85SWE: +16%
Gemini 3.7 Flash vs Qwen3 235BAlibaba
Compare

Open-weight Qwen3 mixture-of-experts.

Elo: +90SWE: +13%
Gemini 3.7 Flash vs Claude Haiku 4.5Anthropic
Compare

Anthropic cheap/fast Claude SKU. Official Claude Platform list price $1/$5 per 1M tokens (Anthropic Haiku page).

Elo: +92SWE: +14.2%
Gemini 3.7 Flash vs Yi-Large01.AI
Compare

01.AI full-scale dense model for complex instruction following.

Elo: +100SWE: +18.6%
Gemini 3.7 Flash vs Qwen 2.5 Coder 32BAlibaba
Compare

Alibaba dedicated open-weight code generation model with near-frontier SWE-bench Verified coding capability.

Elo: +105SWE: +7.2%
Gemini 3.7 Flash vs Kimi Chat 1.5Moonshot
Compare

Moonshot ultra-long context model supporting up to 2 million tokens per request.

Elo: +110SWE: +21.4%
Gemini 3.7 Flash vs Llama 4 ScoutMeta
Compare

Meta open-weight Llama 4 long-context sibling of Maverick. Common public API lists sit near $0.08–$0.30 / $0.30–$0.70 per 1M; we store a conservative hosted list until OpenRouter overwrites.

Elo: +118SWE: +20%
Gemini 3.7 Flash vs GPT-5 miniOpenAI
Compare

Cost-efficient GPT-5 distill for high-volume agents.

Elo: +120SWE: +20%
Gemini 3.7 Flash vs Grok 3xAI
Compare

Previous xAI generation before Grok 4. Kept for grok 3 vs grok 4.

Elo: +125SWE: +21%
Gemini 3.7 Flash vs Qwen3.8 27BAlibaba
Compare

Alibaba open-weight 27B drop dated Aug 14 2026. Dense enough to self-host; not a frontier MoE.

Elo: +132SWE: +13.6%
Gemini 3.7 Flash vs Llama 3.1 405BMeta
Compare

Meta flagship open-weight 405B dense foundation model with 128k context window.

Elo: +160SWE: +14%
Gemini 3.7 Flash vs DeepSeek Coder V2DeepSeek
Compare

DeepSeek open-weight Mixture-of-Experts coding model supporting 338 programming languages and 128k context.

Elo: +165SWE: +11.9%
Gemini 3.7 Flash vs Claude 3.7 SonnetAnthropic
Compare

Hybrid-reasoning Sonnet from 2025. Kept for historical compare pages.

Elo: +168SWE: +2.1%
Gemini 3.7 Flash vs Doubao Lite 1.5ByteDance
Compare

ByteDance high-speed lightweight model priced at sub-cent levels.

Elo: +170SWE: +26.4%
Gemini 3.7 Flash vs DeepSeek R1DeepSeek
Compare

Open-weights reasoning model trained with large-scale RL.

Elo: +172SWE: +7.2%
Gemini 3.7 Flash vs Gemini 2.5 ProGoogle
Compare

Previous Google long-context flagship.

Elo: +180SWE: +8.6%
Gemini 3.7 Flash vs Command ACohere
Compare

Cohere enterprise RAG and tool-use model.

Elo: +186SWE: +29.8%
Gemini 3.7 Flash vs GPT-4oOpenAI
Compare

Previous OpenAI flagship. Still a common compare baseline on legacy pages.

Elo: +195SWE: +17.6%
Gemini 3.7 Flash vs Codestral 25.01Mistral
Compare

Mistral code-specialist model.

Elo: +210SWE: +20.6%
Gemini 3.7 Flash vs DeepSeek V3DeepSeek
Compare

Prior DeepSeek flagship. Baseline for v3 vs v4.

Elo: +220SWE: +23.8%
Gemini 3.7 Flash vs Gemini 2.5 FlashGoogle
Compare

Previous Google speed workhorse.

Elo: +235SWE: +24.4%
Gemini 3.7 Flash vs Llama 3.3 70BMeta
Compare

Previous Meta 70B open-weight workhorse.

Elo: +245SWE: +27.3%
Gemini 3.7 Flash vs Claude 3.5 SonnetAnthropic
Compare

2024 workhorse. Still a high-intent compare against GPT-4o.

Elo: +250SWE: +23.4%
Gemini 3.7 Flash vs GPT-4o miniOpenAI
Compare

Legacy small OpenAI model. Useful as a cheap baseline.

Elo: +258SWE: +31.2%
Gemini 3.7 Flash vs Gemini 1.5 ProGoogle
Compare

First million-token Gemini Pro. Baseline for 1.5 vs 2.5 vs 3.x Pro.

Elo: +270SWE: +34.4%
Gemini 3.7 Flash vs GPT-4 TurboOpenAI
Compare

GPT-4 Turbo 128k. Historical flagship for gpt-4 turbo vs gpt-4o / gpt-5.

Elo: +275SWE: +39.2%
Gemini 3.7 Flash vs Claude 3 OpusAnthropic
Compare

Original Claude 3 flagship. Kept so opus 3 vs later Opus and vs GPT-4o still resolve.

Elo: +282SWE: +34%
Gemini 3.7 Flash vs Llama 3.1 70BMeta
Compare

Llama 3.1 70B instruct. Predecessor to 3.3 70B and Llama 4.

Elo: +290SWE: +32.2%
Gemini 3.7 Flash vs Claude 3.5 HaikuAnthropic
Compare

Previous cheap Claude. Haiku 4.5 is the current $1/$5 SKU.

Elo: +306SWE: +36.2%

Related news

  • compare · Aug 13, 2026

    Gemini Flash vs Pro: when Flash is the whole product

  • price · Aug 13, 2026

    Gemini 3.6 Flash (Jul 21) now shares the 3.7 intro rate

  • launch · Aug 13, 2026

    Gemini 3.7 Flash (Aug 13): $0.75/$3.75 intro through Dec 31 2026

  • launch · Jul 15, 2026

    Gemini 3.6 Pro: Google’s current Pro-class long-context row

Frequently asked questions

Plain-English methodology and leaderboard answers

Gemini 3.7 Flash is a Google closed-API model. Google Aug 13 2026 workhorse. Official intro price $0.75/$3.75 per 1M through Dec 31 2026; 1,048,576-token context (Google blog).

Elo is a crowd vote on which hidden answer people liked more — not a school test. What is Elo?

Where this model ranks in Stack Engine

cheap coding agents

SWE-bench first, then price. For CI bots and repo agents that cannot burn Opus prices.

#2

lowest-latency chat

TTFT and tokens/sec first. For support widgets and voice-adjacent loops.

#1

long-context RAG

Context window and input price for stuffing large corpora.

#3

frontier agents

Coding + preference Elo for computer-use and multi-step tools.

#1

vision and screenshots

Multimodal models only. Preference Elo and latency for UI-understanding jobs.

#1

cheapest hosted API

Output price first among models we still consider usable for chat.

#1

writing and editing

Preference Elo first for long-form drafts. Price still counts if you generate all day.

#1

mid-tier workhorse

Sonnet / Terra / Flash class. Usable Elo without Opus or Sol prices.

#1