July 15, 2026
Three providers offer fourteen models with a 200x price gap between cheapest and most expensive. This report breaks down every tier so teams can pick the right model.
The flagship tier holds steady at $5/MTok input for Claude Opus 4.8 and GPT 5.5. GPT 5.6 Sol matches that $5/MTok price point at launch this quarter.
Claude Fable 5 sits at the premium end: $10/MTok input and $50/MTok output. That makes Fable 5 the most expensive model in the entire registry today.
Output pricing reveals a separate gap worth tracking across providers and workload types. GPT 5.5 and GPT 5.6 Sol charge $30/MTok output, versus Claude Opus at $25/MTok. For output-heavy workloads like code generation, that 20% gap compounds over thousands of calls.
Gemini 3.1 Pro is the cheapest mid-tier option at $2.00 per million input tokens. Claude Sonnet 5 introductory pricing makes it temporarily competitive at $2/$10 through August 2026.
Gemini 3.5 Flash-Lite at $0.25/MTok input is the cheapest model across all three providers. For classification and extraction tasks it scores 85/100 confidence at a fraction of flagship cost.
Run your numbers through our free teardown at /teardown, or start tracking real spend with a free dashboard.
Calculators estimate. The dashboard shows what you actually spend and where you can save.