CompareLLM
CompareLLM.ai
Live
Leaderboard
Quality & Reasoning
Overall Arena EloPrimary

LMSYS crowd human preference ranking

Coding Elo & SWE-benchCode

Real GitHub issue software solve rate

LiveBench Reasoning

Contamination-free automated tests

GPQA Diamond

PhD-level science & domain knowledge

Speed & Token Cost
Throughput (tok/s)Speed

Output generation token rate

Time to First Token (TTFT)

Response latency for voice & chat loops

Output Price ($/1M tokens)

Cost per million generated tokens

Live Pareto Frontier Scatter

Quality vs Cost efficiency boundary

Models
Model Classes
Frontier ModelsProprietary

Opus 4.5, GPT-5, Gemini 2.0 Pro

Open-Weight CatalogApache/MIT

Llama 3.3, DeepSeek, Qwen 2.5

🇨🇳 Chinese LLMsCN

DeepSeek V3, Qwen, GLM-5, MiniMax

Browse All 40+ Models
Top Providers
Anthropic

Claude Opus 4.5, Sonnet 4.5, Haiku

OpenAI

GPT-5, GPT-4.5, GPT-4o, o3

DeepSeek

DeepSeek V3, R1 Reasoning

Google

Gemini 2.0 Pro, Flash, Thinking

Compare
Popular Head-to-Head ShowdownsView all 48+ pairs →
🇨🇳 DeepSeek V3 vs 🇺🇸 GPT-5

East vs West frontier battle

Sonnet 4.5 vs 🇨🇳 DeepSeek V3

Everyday developer favorite

🇨🇳 Qwen 2.5 vs 🇺🇸 Llama 3.3

Open-weights value clash

Claude Opus 4.5 vs GPT-5

Flagship proprietary duel

Open Interactive Comparison Matrix
Best of & Stacks
Best LLM Lists (2026)
Best Coding LLM

SWE-bench verified repository tests

Best Cheap LLM

Sub-$1/1M token value powerhouses

Fastest Low-Latency LLM

Sub-200ms TTFT for voice & live chat

Claude vs GPT Benchmark

Anthropic vs OpenAI head-to-head

Stack Engine Presets
Cheap Coding Agents

Budget repo bots with high SWE-bench

Lowest-Latency Chat

Fast interactive support loops

Open-Weight Reasoning

Self-hostable reasoning power

Explore All 11 Presets
Research & News
Intelligence & Telemetry
News & Benchmark BriefingsDispatches

Verified model promotions & price shifts

Hourly Benchmark ChangelogLive

Dated ingest audit trail with exact diffs

Evaluation Methodology

Standardized scoring formulas & harnesses

What is Arena Elo?

Understanding blind pairwise human ratings

…
CompareLLM
CompareLLM.ai
Precision Benchmarks

Programmatic, dated AI model benchmarks, head-to-head comparisons, and Stack Engine presets.

Daily ingest · 06:00 UTC

Analytics & Benchmarks

  • AI Model Leaderboard
  • Head-to-Head Compare Hub
  • Models Directory
  • Stack Engine Presets
  • Frontier Models
  • Open Weights Catalog

Guides & Intent Lists

  • Best LLM Lists (2026)
  • Best Coding LLM
  • Best Cheap LLM
  • Fastest Low-Latency LLM
  • Claude vs GPT Benchmark
  • What is Elo?
  • Methodology Guides
  • News & Dispatches

Transparency & API

  • Evaluation Methodology
  • Benchmark Changelog
  • Public JSON API
  • llms.txt Specification
  • Privacy Policy
  • Sign In / Account

© 2026 CompareLLM. Public benchmark data aggregated from Arena Elo, LiveBench, SWE-bench & OpenRouter.

Every score has a dated snapshot.

  1. Home
  2. Models
  3. Recraft V4.1 Utility Pro
RecraftFrontier Model

Recraft V4.1 Utility Pro

Recraft V4.1 Utility Pro is a Recraft closed-API frontier model. Specialized generative model for vector art, icon sets, illustrations, and commercial design assets with native brand color matching. Numbers below are dated snapshots, not a guarantee on your traffic mix.

At a glance

Recraft V4.1 Utility Pro is a Recraft closed-API frontier model. Specialized generative model for vector art, icon sets, illustrations, and commercial design assets with native brand color matching. Numbers below are dated snapshots, not a guarantee on your traffic mix.
Last updated Aug 1, 2026
All 100 vsSave Model

Compare Recraft V4.1 Utility Pro against

Capability Profile

Recraft V4.1 Utility Pro Benchmark Percentiles

Plotted against all active catalog models (50th percentile = catalog median).

Recraft V4.1 Utility Pro
Median (50)

Dimensional Scorecards

Visual Quality
26%
Prompt Adherence
64%
Typography
84%
Gen Speed
22%
Price Value
5%
Standout Competencies

Ranks in the top tier (≥75th percentile) for Typography.

Benchmark Specifications

Dated snapshot metrics aggregated from official evaluators and API providers.

Image Elo1,218
seed-bootstrap · Aug 1, 2026
Generation time4.6s
seed-bootstrap · Aug 1, 2026
Price per 1k images$210/1k
seed-bootstrap · Aug 1, 2026
Prompt adherence91%
seed-bootstrap · Aug 1, 2026
Text rendering94%
seed-bootstrap · Aug 1, 2026
Benchmark MetricReported ScoreObserved Source
Image Elo1,218seed-bootstrap · Aug 1, 2026
Generation time4.6sseed-bootstrap · Aug 1, 2026
Price per 1k images$210/1kseed-bootstrap · Aug 1, 2026
Prompt adherence91%seed-bootstrap · Aug 1, 2026
Text rendering94%seed-bootstrap · Aug 1, 2026

Compare with every model

Search or pick any catalog row. Suggested matchups first, then the full list.

Dedicated vs hub →
Recraft V4.1 Utility Pro vs Recraft V4.1 Utility (previous recraft)Recraft V4.1 Utility Pro vs GPT Image 2 (High)Recraft V4.1 Utility Pro vs GPT Image 2 (Low / Fast)Recraft V4.1 Utility Pro vs GPT Image 1.5Recraft V4.1 Utility Pro vs DALL-E 3Recraft V4.1 Utility Pro vs Reve 2.1
Showing 100 of 100 comparisons
Recraft V4.1 Utility Pro vs Claude Opus 5Anthropic
Compare

Anthropic current default flagship. Official API $5/$25 per 1M tokens and a 1M context window (Anthropic, Jul 24 2026).

Recraft V4.1 Utility Pro vs Claude Fable 5Anthropic
Compare

Anthropic top-tier long-horizon model. Official API $10/$50 per 1M tokens (Claude Platform pricing, Aug 2026).

Recraft V4.1 Utility Pro vs GPT-5.6 SolOpenAI
Compare

OpenAI 5.6 flagship tier. Official API $5/$30 per 1M tokens (OpenAI pricing, Jul 30 2026 update left Sol unchanged).

Recraft V4.1 Utility Pro vs Claude Opus 4.8Anthropic
Compare

Prior Opus generation still billed at $5/$25. Kept as a compare baseline against Opus 5.

Recraft V4.1 Utility Pro vs Grok 4.6xAI
Compare

xAI Aug 12 2026 post-training refresh of Grok 4.5. Same $2/$6 API price, 500k context, stronger agentic traces.

Recraft V4.1 Utility Pro vs Claude Opus 4.6Anthropic
Compare

Follow-on Opus release. Slightly behind 4.5 on official SWE-bench bash-only in the last published sweep; stronger long-horizon agent traces.

Recraft V4.1 Utility Pro vs Gemini 3.6 ProGoogle
Compare

Current Google Pro-class multimodal model. Long context, strong coding, billed like the 3.x Pro tier.

Recraft V4.1 Utility Pro vs Claude Opus 4.5Anthropic
Compare

Anthropic frontier coding and computer-use model. SWE-bench leader on the official mini-SWE-agent harness in the Feb 2026 refresh.

Recraft V4.1 Utility Pro vs OpenAI: o3 MiniOpenAI
Compare

Auto-discovered from OpenRouter (openai/o3-mini). Preview until a second source matches.

Recraft V4.1 Utility Pro vs GPT-5OpenAI
Compare

OpenAI flagship reasoning model for 2025–26. Strong general preference Elo and multimodal coverage.

Recraft V4.1 Utility Pro vs GLM-5.3Zhipu
Compare

Z.ai Aug 14 2026 post-train of the GLM-5.2 744B base. Coding-plan live; open weights promised after a two-week safety review (z.ai/blog/glm-5.3).

Recraft V4.1 Utility Pro vs Claude Sonnet 5Anthropic
Compare

Anthropic workhorse. Official $2/$10 per 1M tokens made permanent on Aug 10 2026.

Recraft V4.1 Utility Pro vs Gemini 3 ProGoogle
Compare

Google frontier multimodal model with a multi-million-token context window.

Recraft V4.1 Utility Pro vs GPT-5.6 TerraOpenAI
Compare

OpenAI 5.6 mid tier. Official API $2/$12 per 1M after the Jul 30 2026 price cut.

Recraft V4.1 Utility Pro vs GPT-4.5 OrionOpenAI
Compare

OpenAI largest dense non-reasoning model with expansive world knowledge and reduced hallucinations.

Recraft V4.1 Utility Pro vs DeepSeek V4 ProDeepSeek
Compare

Open-weight-adjacent DeepSeek flagship. High reasoning density per dollar.

Recraft V4.1 Utility Pro vs Gemini 3.7 FlashGoogle
Compare

Google Aug 13 2026 workhorse. Official intro price $0.75/$3.75 per 1M through Dec 31 2026; 1,048,576-token context (Google blog).

Recraft V4.1 Utility Pro vs Kimi K3Moonshot
Compare

Moonshot AI frontier flagship reasoning and agent swarm model with 256k context and top-tier SWE-bench coding capability.

Recraft V4.1 Utility Pro vs Claude Sonnet 4.5Anthropic
Compare

Workhorse Anthropic model: most of Opus coding quality at a mid-tier price.

Recraft V4.1 Utility Pro vs Qwen 3 MaxAlibaba
Compare

Alibaba flagship. Strong math and multilingual code.

Recraft V4.1 Utility Pro vs MoonshotAI: Kimi K2.5Moonshot
Compare

Auto-discovered from OpenRouter (moonshotai/kimi-k2.5). Preview until a second source matches.

Recraft V4.1 Utility Pro vs OpenAI: o1OpenAI
Compare

Auto-discovered from OpenRouter (openai/o1). Preview until a second source matches.

Recraft V4.1 Utility Pro vs Gemini 3.6 FlashGoogle
Compare

Google Jul 21 workhorse. Now shares the 3.7 Flash introductory $0.75/$3.75 rate through Dec 31 2026.

Recraft V4.1 Utility Pro vs Gemini 3 FlashGoogle
Compare

Fast Google frontier-adjacent model. Near-top official SWE-bench at a fraction of Opus price.

Recraft V4.1 Utility Pro vs Qwen QwQ 32BAlibaba
Compare

Alibaba specialized open reasoning model competing with frontier closed reasoning models.

Recraft V4.1 Utility Pro vs GLM-5.2Zhipu
Compare

Zhipu flagship. Strong Chinese/English coding and agents.

Recraft V4.1 Utility Pro vs Gemini 2.0 Flash ThinkingGoogle
Compare

Google experimental reasoning model that visualizes thoughts in real-time.

Recraft V4.1 Utility Pro vs DeepSeek V4 FlashDeepSeek
Compare

DeepSeek Jul 31 2026 price-performance SKU. Public listings put it near $0.14/$0.28 per 1M tokens.

Recraft V4.1 Utility Pro vs Claude Opus 4Anthropic
Compare

First Claude 4 Opus generation. Baseline for opus 4 vs opus 5.

Recraft V4.1 Utility Pro vs Grok 4xAI
Compare

Previous xAI flagship.

Recraft V4.1 Utility Pro vs MiniMax M2.5MiniMax
Compare

MiniMax coding model. Tied near the top of official SWE-bench bash-only in Feb 2026.

Recraft V4.1 Utility Pro vs Yi-Lightning01.AI
Compare

01.AI ultra-fast reasoning model delivering top LiveBench efficiency.

Recraft V4.1 Utility Pro vs Z.ai: GLM 5V TurboZhipu
Compare

Auto-discovered from OpenRouter (z-ai/glm-5v-turbo). Preview until a second source matches.

Recraft V4.1 Utility Pro vs Seed 2.1 TurboByteDance
Compare

ByteDance Seed 2.1 Turbo, listed on public model timelines as an Aug 10 2026 API drop.

Recraft V4.1 Utility Pro vs Doubao Pro 1.5ByteDance
Compare

ByteDance flagship enterprise model with ultra-low token cost and 128k context.

Recraft V4.1 Utility Pro vs Llama 4 MaverickMeta
Compare

Meta natively multimodal open-weight flagship.

Recraft V4.1 Utility Pro vs GPT-5.6 LunaOpenAI
Compare

OpenAI 5.6 fast/cheap tier. Official API $0.20/$1.20 per 1M after the Jul 30 80% Luna cut.

Recraft V4.1 Utility Pro vs MoonshotAI: Kimi K2 0711Moonshot
Compare

Auto-discovered from OpenRouter (moonshotai/kimi-k2). Preview until a second source matches.

Recraft V4.1 Utility Pro vs Mistral Large 3Mistral
Compare

Mistral European flagship with strong function calling.

Recraft V4.1 Utility Pro vs Claude Sonnet 4Anthropic
Compare

First Sonnet 4 generation. Bridge between 3.5/3.7 and Sonnet 5.

Recraft V4.1 Utility Pro vs Ernie 4.5 TurboBaidu
Compare

Baidu current multimodal enterprise foundation model with broad Chinese knowledge.

Recraft V4.1 Utility Pro vs Qwen 2.5 PlusAlibaba
Compare

Alibaba balanced flagship API model with high-throughput general reasoning.

Recraft V4.1 Utility Pro vs OpenAI o1-miniOpenAI
Compare

OpenAI high-speed, cost-effective reasoning model optimized for STEM, math, and code generation.

Recraft V4.1 Utility Pro vs Qwen3 235BAlibaba
Compare

Open-weight Qwen3 mixture-of-experts.

Recraft V4.1 Utility Pro vs Claude Haiku 4.5Anthropic
Compare

Anthropic cheap/fast Claude SKU. Official Claude Platform list price $1/$5 per 1M tokens (Anthropic Haiku page).

Recraft V4.1 Utility Pro vs Yi-Large01.AI
Compare

01.AI full-scale dense model for complex instruction following.

Recraft V4.1 Utility Pro vs Qwen 2.5 Coder 32BAlibaba
Compare

Alibaba dedicated open-weight code generation model with near-frontier SWE-bench Verified coding capability.

Recraft V4.1 Utility Pro vs Qwen2.5 Coder 32B InstructAlibaba
Compare

Auto-discovered from OpenRouter (qwen/qwen-2.5-coder-32b-instruct). Preview until a second source matches.

Recraft V4.1 Utility Pro vs Kimi Chat 1.5Moonshot
Compare

Moonshot ultra-long context model supporting up to 2 million tokens per request.

Recraft V4.1 Utility Pro vs Llama 4 ScoutMeta
Compare

Meta open-weight Llama 4 long-context sibling of Maverick. Common public API lists sit near $0.08–$0.30 / $0.30–$0.70 per 1M; we store a conservative hosted list until OpenRouter overwrites.

Recraft V4.1 Utility Pro vs GPT-5 miniOpenAI
Compare

Cost-efficient GPT-5 distill for high-volume agents.

Recraft V4.1 Utility Pro vs Grok 3xAI
Compare

Previous xAI generation before Grok 4. Kept for grok 3 vs grok 4.

Recraft V4.1 Utility Pro vs Qwen3.8 27BAlibaba
Compare

Alibaba open-weight 27B drop dated Aug 14 2026. Dense enough to self-host; not a frontier MoE.

Recraft V4.1 Utility Pro vs Llama 3.1 405BMeta
Compare

Meta flagship open-weight 405B dense foundation model with 128k context window.

Recraft V4.1 Utility Pro vs DeepSeek Coder V2DeepSeek
Compare

DeepSeek open-weight Mixture-of-Experts coding model supporting 338 programming languages and 128k context.

Recraft V4.1 Utility Pro vs Claude 3.7 SonnetAnthropic
Compare

Hybrid-reasoning Sonnet from 2025. Kept for historical compare pages.

Recraft V4.1 Utility Pro vs Doubao Lite 1.5ByteDance
Compare

ByteDance high-speed lightweight model priced at sub-cent levels.

Recraft V4.1 Utility Pro vs DeepSeek R1DeepSeek
Compare

Open-weights reasoning model trained with large-scale RL.

Recraft V4.1 Utility Pro vs Gemini 2.5 ProGoogle
Compare

Previous Google long-context flagship.

Recraft V4.1 Utility Pro vs Command ACohere
Compare

Cohere enterprise RAG and tool-use model.

Recraft V4.1 Utility Pro vs GPT-4oOpenAI
Compare

Previous OpenAI flagship. Still a common compare baseline on legacy pages.

Recraft V4.1 Utility Pro vs Codestral 25.01Mistral
Compare

Mistral code-specialist model.

Recraft V4.1 Utility Pro vs DeepSeek V3DeepSeek
Compare

Prior DeepSeek flagship. Baseline for v3 vs v4.

Recraft V4.1 Utility Pro vs Gemini 2.5 FlashGoogle
Compare

Previous Google speed workhorse.

Recraft V4.1 Utility Pro vs Llama 3.3 70BMeta
Compare

Previous Meta 70B open-weight workhorse.

Recraft V4.1 Utility Pro vs Claude 3.5 SonnetAnthropic
Compare

2024 workhorse. Still a high-intent compare against GPT-4o.

Recraft V4.1 Utility Pro vs GPT-4o miniOpenAI
Compare

Legacy small OpenAI model. Useful as a cheap baseline.

Recraft V4.1 Utility Pro vs Gemini 1.5 ProGoogle
Compare

First million-token Gemini Pro. Baseline for 1.5 vs 2.5 vs 3.x Pro.

Recraft V4.1 Utility Pro vs GPT-4 TurboOpenAI
Compare

GPT-4 Turbo 128k. Historical flagship for gpt-4 turbo vs gpt-4o / gpt-5.

Recraft V4.1 Utility Pro vs Claude 3 OpusAnthropic
Compare

Original Claude 3 flagship. Kept so opus 3 vs later Opus and vs GPT-4o still resolve.

Recraft V4.1 Utility Pro vs Llama 3.1 70BMeta
Compare

Llama 3.1 70B instruct. Predecessor to 3.3 70B and Llama 4.

Recraft V4.1 Utility Pro vs Claude 3.5 HaikuAnthropic
Compare

Previous cheap Claude. Haiku 4.5 is the current $1/$5 SKU.

Recraft V4.1 Utility Pro vs GPT Image 2 (High)OpenAI
Compare

OpenAI flagship diffusion-transformer image synthesis model with supreme prompt adherence, realistic textures, and complex text composition.

Recraft V4.1 Utility Pro vs GPT Image 2 (Low / Fast)OpenAI
Compare

Fast distilled tier of GPT Image 2 designed for high-throughput interactive creative workflows at 70% lower price.

Recraft V4.1 Utility Pro vs GPT Image 1.5OpenAI
Compare

Previous OpenAI image generation flagship. High fidelity with proven enterprise reliability.

Recraft V4.1 Utility Pro vs DALL-E 3OpenAI
Compare

Legacy OpenAI image model benchmark baseline. Retained for historical comparisons.

Recraft V4.1 Utility Pro vs Reve 2.1Reve
Compare

State-of-the-art cinematic image synthesis foundation model known for photorealistic lighting and aesthetic composition.

Recraft V4.1 Utility Pro vs Nano Banana Pro (Gemini 3 Pro Image)Google
Compare

Google DeepMind frontier multimodal image generation flagship with Deep Research and complex multi-object spatial reasoning.

Recraft V4.1 Utility Pro vs Nano Banana 2 (Gemini 3.1 Flash Image)Google
Compare

Google high-speed multimodal generative model delivering top Elo performance at half the latency and cost of Pro.

Recraft V4.1 Utility Pro vs Nano Banana 2 LiteGoogle
Compare

Sub-1.5 second ultra-low latency tier for real-time applications and mobile game asset pipelines.

Recraft V4.1 Utility Pro vs FLUX.2 [max]Black Forest Labs
Compare

Black Forest Labs maximum-capacity flow-matching model with photoreal anatomy and leading typography fidelity.

Recraft V4.1 Utility Pro vs FLUX.2 [flex]Black Forest Labs
Compare

Flexible step-distilled variant of FLUX.2 offering 99% of Max quality with 45% faster generation speeds.

Recraft V4.1 Utility Pro vs FLUX.1.1 [pro]Black Forest Labs
Compare

First-generation professional API model from Black Forest Labs, renowned for 6x faster generation than original FLUX.1 Pro.

Recraft V4.1 Utility Pro vs FLUX.1 [dev]Black Forest Labs
Compare

Open-weights non-commercial base model with 12B parameters, serving as the foundation for the open-source fine-tuning ecosystem.

Recraft V4.1 Utility Pro vs FLUX.1 [schnell]Black Forest Labs
Compare

Apache 2.0 4-step distilled open weights model optimized for local inference and instant preview generation.

Recraft V4.1 Utility Pro vs Ideogram 4.0 (Quality)Ideogram
Compare

Industry-leading typography, graphic design, and in-image text layout model with precise kerning and complex banner generation.

Recraft V4.1 Utility Pro vs Ideogram 4.0 Open WeightsIdeogram
Compare

Open-weight release of Ideogram 4.0, bringing world-class text rendering to self-hosted enterprise infrastructure.

Recraft V4.1 Utility Pro vs Recraft V4.1 UtilityRecraft
Compare

High-speed utility tier of Recraft V4.1 tailored for rapid asset production at $0.035/image.

Recraft V4.1 Utility Pro vs Krea 2 LargeKrea
Compare

High-resolution real-time generation model tuned for artistic composition and dynamic concept art.

Recraft V4.1 Utility Pro vs Krea 2 Medium TurboKrea
Compare

Sub-2 second interactive generation engine delivering ultra-affordable $15/1k image generation.

Recraft V4.1 Utility Pro vs Qwen Image 2.0 ProAlibaba
Compare

Alibaba multimodal image synthesis foundation model with strong bilingual Chinese/English typography and cultural asset fidelity.

Recraft V4.1 Utility Pro vs Wan2.6 Text to ImageAlibaba
Compare

Alibaba high-efficiency open video/image unified transformer backbone.

Recraft V4.1 Utility Pro vs Seedream 5.0 ProByteDance
Compare

ByteDance flagship image generation model with high photorealism and fine facial structure rendering.

Recraft V4.1 Utility Pro vs Seedream 4.0ByteDance
Compare

High-value ByteDance image foundation model at $30/1k images.

Recraft V4.1 Utility Pro vs Midjourney v7Midjourney
Compare

Midjourney v7 premier creative image generation model with unparalleled stylistic nuances, coherent hands/anatomy, and cinematic grading.

Recraft V4.1 Utility Pro vs Midjourney v6.1Midjourney
Compare

Previous standard in aesthetic digital art generation, retained as a historical compare baseline.

Recraft V4.1 Utility Pro vs MAI-Image-2.5Microsoft AI
Compare

Microsoft AI proprietary foundation image model for Copilot Studio and enterprise creative workflows.

Recraft V4.1 Utility Pro vs MAI-Image-2.5-FlashMicrosoft AI
Compare

Cost-optimized Microsoft AI image model delivering sub-2 second responses at $20/1k images.

Recraft V4.1 Utility Pro vs HiDream-O1-Image-1.5HiDream
Compare

HiDream high-definition visual generator optimized for photorealism and accurate complex lighting.

Recraft V4.1 Utility Pro vs Luma UNI 1 MaxLuma Labs
Compare

Luma Labs universal 3D-aware image synthesis engine with spatial geometry consistency.

Frequently asked questions

Plain-English methodology and leaderboard answers

Recraft V4.1 Utility Pro is a Recraft closed-API model. Specialized generative model for vector art, icon sets, illustrations, and commercial design assets with native brand color matching.

Elo is a crowd vote on which hidden answer people liked more — not a school test. What is Elo?