Every model, one page each
23 models from 11 providers. Each page collects the benchmark scores, prices, and context window for one model, plus who it suits and what it trades away. Want them side by side instead? Compare them all.
Anthropic
- Anthropic's flagship for agentic coding and long business workflows. It comes close to Fable 5 at half the price.$5 in · $25 out · 1M context · 6 benchmarks
Claude Opus 5
- Anthropic's most capable model. Built for the hardest reasoning and long autonomous work, at a premium price.$10 in · $50 out · 1M context · 7 benchmarks
Claude Fable 5
- A coding workhorse. Near the top of the toughest coding benchmarks at half the price of Fable 5.$5 in · $25 out · 1M context · 6 benchmarks
Claude Opus 4.8
- The most agentic Sonnet yet, released June 30. Can make plans and run tools autonomously. Best value for daily work, now at intro pricing.$2 in · $10 out · 1M context · 7 benchmarks
Claude Sonnet 5
- Anthropic's fastest and cheapest model. Great for quick answers and simple tasks in high volume.$1 in · $5 out · 200K context · 2 benchmarks
Claude Haiku 4.5
OpenAI
- OpenAI's brand-new flagship. State of the art on autonomous terminal work, strong all-rounder.$5 in · $30 out · 1.05M context · 7 benchmarks
GPT-5.6 Sol
- The balanced middle tier of the GPT-5.6 family. Most of Sol's ability at half the cost.$2.50 in · $15 out · 1.05M context · 6 benchmarks
GPT-5.6 Terra
- The fastest, most cost-efficient GPT-5.6 tier. Built for speed and high-volume simple tasks.$1 in · $6 out · 1.05M context · 7 benchmarks
GPT-5.6 Luna
- Google's flagship. A top-tier reasoner with strong long-context skills at an aggressive price.$2 in · $12 out · 1M context · 7 benchmarks
Gemini 3.1 Pro
- Google's newest workhorse model. Stronger at agentic coding than 3.5 Flash while using fewer tokens and charging less for output.$1.50 in · $7.50 out · 1.05M context · 4 benchmarks
Gemini 3.6 Flash
- Google's fastest, lowest-cost 3.5 model. Built for high-volume extraction, analysis, and autonomous subagent work.$0.30 in · $2.50 out · 1.05M context · 3 benchmarks
Gemini 3.5 Flash-Lite
- Google's speed tier that punches above its weight. Beats 3.1 Pro on coding at about 25% lower cost.$1.50 in · $9 out · 1M context · 6 benchmarks
Gemini 3.5 Flash
xAI
- xAI's brand-new flagship. Strong terminal and coding chops at a mid-tier price, with live access to X (Twitter) data.$2 in · $6 out · 500K context · 7 benchmarks
Grok 4.5
- A budget speedster with a huge 2M-token context window. One of the cheapest ways to process large amounts of text.$0.20 in · $0.50 out · 2M context · 3 benchmarks
Grok 4.1 Fast
Meta
- Meta's new flagship and its first paid, closed-weights model after the open Llama era. Built for agent work at an aggressive price.$1.25 in · $4.25 out · 1M context · 6 benchmarks
Muse Spark 1.1
- Meta's general-purpose open model. Easy to run and widely supported, though newer open models beat it on hard reasoning.Self-hosted pricing · 1M context · 2 benchmarks
Llama 4 Maverick
Open source - The long-context champion. A 10-million-token window, enough to read hundreds of books at once.Self-hosted pricing · 10M context · 2 benchmarks
Llama 4 Scout
Open source