AI Stack Engine Presets
Deterministic recommendations computed from verified benchmark percentiles. Objective weights published transparently on each page.
cheap coding agents
SWE-bench first, then price. For CI bots and repo agents that cannot burn Opus prices.
open-weight reasoning
Highest reasoning signal among models you can self-host or buy as open weights.
lowest-latency chat
TTFT and tokens/sec first. For support widgets and voice-adjacent loops.
long-context RAG
Context window and input price for stuffing large corpora.
frontier agents
Coding + preference Elo for computer-use and multi-step tools.
vision and screenshots
Multimodal models only. Preference Elo and latency for UI-understanding jobs.
cheapest hosted API
Output price first among models we still consider usable for chat.
highest coding accuracy
SWE-bench first with no budget cap. For when the patch quality matters more than the invoice.
writing and editing
Preference Elo first for long-form drafts. Price still counts if you generate all day.
mid-tier workhorse
Sonnet / Terra / Flash class. Usable Elo without Opus or Sol prices.
cheapest frontier
Models that still clear a high Elo bar, sorted so price hurts.
