August 16, 2026
Here is every Claude, GPT and Gemini model we track, sorted by input price per million tokens. Prices come straight from the model registry, so this table is current rather than a snapshot.
Fast tier ($0.20 to $1.50/MTok input):
Mid tier ($2.00 to $3.00/MTok input):
Flagship tier ($4.00 to $10.00/MTok input):
xAI is the fourth provider we track, and its Grok family is deliberately not in the tables above - this comparison is scoped to the three providers in the title. Grok prices are on the model pages and in the teardown; we publish no quality scores for them yet, because our weekly benchmark source does not cover Grok.
The cheapest model per task type depends on the minimum quality threshold you need. At 80+ confidence score:
The budget end is not one provider's territory any more. GPT 5.6 Luna at $0.20/MTok now undercuts every Gemini model after OpenAI's August cuts, and it clears the 80-confidence bar on all four high-volume tasks. In the mid tier the three providers land on the same $2.00/MTok list price, so the choice there is a quality call, not a price one.
OpenAI and Google price cached input at 0.10x list - 90% off - on every model, and Anthropic does on all but two (Claude Fable 5.1 discounts further still, at 98%). What still differs is the cache-WRITE fee: 25% on Anthropic, 25% on the GPT 5.6 family, nothing on GPT 5.5, GPT 5.4 or any Gemini model.
At an 80% hit rate, with the misses paying the write fee, a million prompt tokens costs:
Because the discount is identical, caching does not reorder models that differ on list price. It does split models that tie on it: Claude Opus 4.8 and GPT 5.5 both list at $5.00, and once you are paying to write cache, GPT 5.5 is the cheaper prompt. Pick on list price and output price, then check the write fee if your prefix churns.
For output-heavy workloads (code generation, creative writing, blog writing), output tokens often outnumber input tokens 2-4x. Output pricing determines total cost:
For code generation, Claude Sonnet 5 is the best value in the mid tier, and no longer a temporary one: its $2.00 / $10.00 launch rate became the standard price when Anthropic cancelled the increase that had been scheduled for September 2026.
No single provider is cheapest across all workloads. The optimal strategy is routing: send classification to the cheapest model that clears your quality bar, code review to a mid-tier model, and only genuinely hard reasoning to a flagship. A team that routes by task type instead of defaulting to one model saves 40-60% on their total AI bill.
Compare all models for your specific workload at /teardown, or start tracking your multi-provider spend with the free dashboard.
Three concrete techniques that teams use to spend less on Claude Code while writing more code. Model-tier picking, prompt caching, and prompt-length discipline.
Provider dashboards show totals, not decisions. Here's how to connect every dollar to the developer and query that spent it.
Real usage data shows Claude Code costs $50-$400 per developer per month depending on model mix and prompt habits. Here is how to estimate and reduce your team's spend.
Calculators estimate. The dashboard shows what you actually spend and where you can save.