7 models across 3 tiers
Current rates for all 7 Claude models, including cache read/write, batch discounts, and long-context tiers.
Models
7
Cheapest input
$1.00/MTok
Cheapest output
$5.00/MTok
Cache discount
90% - 98% off
| Model | Tier | Input $/MTok | Output $/MTok | Cache Read | Cache Write | Batch | Context |
|---|---|---|---|---|---|---|---|
| Claude Haiku 4.5 | Fast / Budget | $1.00 | $5.00 | $0.10(90% off) | $1.25(25% surcharge) | $0.50(50% off) | 200K |
| Claude Sonnet 5 | Mid-tier | $2.00 | $10.00 | $0.20(90% off) | $2.50(25% surcharge) | $1.00(50% off) | 200K |
| Claude Sonnet 4.6 | Mid-tier | $3.00 | $15.00 | $0.30(90% off) | $3.75(25% surcharge) | $1.50(50% off) | 200K |
| Claude Opus 4.8 | Flagship | $5.00 | $25.00 | $0.50(90% off) | $6.25(25% surcharge) | $2.50(50% off) | 200K |
| Claude Opus 5 | Flagship | $5.00 | $25.00 | $0.50(90% off) | $6.25(25% surcharge) | $2.50(50% off) | 1,000K |
| Claude Fable 5.1 | Flagship | $10.00 | $50.00 | $0.25(98% off) | $12.50(25% surcharge) | $5.00(50% off) | 1,000K |
| Claude Fable 5 | Flagship | $10.00 | $50.00 | $1.00(90% off) | $12.50(25% surcharge) | $5.00(50% off) | 1,000K |
Claude Sonnet 5: Launched at $2/$10 as introductory pricing through Aug 31, 2026. Anthropic cancelled the scheduled increase to $3/$15, so this is now the standard price.
How caching and batch processing affect your Claude API costs.
When a prompt prefix is already cached, the provider charges a reduced rate for those tokens. Claude cache read discounts range from 90% to 98% off the standard input rate, depending on the model.
All Claude models charge a 25% surcharge on cache writes. The first time tokens enter cache, they cost 125% of the standard input rate.
All Claude models offer a 50% batch discount. Batch requests are processed asynchronously and cost 50% of the standard rate.
Estimated cost per API call across common workloads, using default token counts for each use case.
Claude Sonnet 5 input decreased
$3.00 to $2.00/MTok
Claude Sonnet 5 output decreased
$15.00 to $10.00/MTok
Claude Sonnet 5 input launched
$3.00/MTok at launch
Claude Fable 5 input launched
$10.00/MTok at launch
Claude Opus 4.8 input increased
$4.50 to $5.00/MTok
5 models from $0.20/MTok
4 models from $0.30/MTok
4 models from $1.00/MTok
See exactly how much you spend on Claude models and where you can save. Two-line wrapper install, no API keys shared.
Claude API pricing starts at $1.00/MTok for input and $5.00/MTok for output. Prices vary by model tier: flagship models cost more but deliver higher quality, while fast/budget models are significantly cheaper for high-volume workloads.
Yes. Claude supports batch processing, which sends requests asynchronously at a reduced rate. The batch discount is 50% off standard pricing across all Claude models.
Claude cache read discounts vary by model, ranging from 90% to 98% off the standard input rate.
There are currently 7 Claude models available through the API, spanning 3 tiers: Fast / Budget, Mid-tier, Flagship.
Claude API pricing is pay-per-use based on token consumption. There is no permanent free tier, but new accounts typically receive introductory credits. Tokeven helps you track exactly what you spend so nothing is wasted.