GPT/Mid-tier

GPT 4o

GPT 4o is GPT's balanced performance model, priced at $2.50/MTok input and $10.00/MTok output with a 128K context window. Backfilled on 2026-09-23. Prices confirmed by the provider pricing page and the LiteLLM cross-check. Release date not published in a machine-readable source. Not yet benchmarked, so never recommended.

Input $/MTok

$2.50

Output $/MTok

$10.00

Context window

128K

Cache read discount

50% off

GPT 4o pricing breakdown

RateInput $/MTokOutput $/MTok
Standard$2.50$10.00
Cached read (50% off input)$1.25-
Cached write (no surcharge)$2.50-
Batch (50% off)$1.25$5.00

All rates in USD per million tokens. Prices verified against provider documentation.

GPT 4o task benchmarks

Confidence scores (0-100) and cross-provider rank for each task type. Rank 1 is best across all 72 models.

Not benchmarked. The third-party benchmark source used for confidence scores does not cover GPT models yet. Scores will appear here once independent benchmarks are published.

GPT 4o cost by use case

Estimated cost per API call using default token counts for each workload.

Use caseInput tokensOutput tokensCost / call
Code Generation2,0004,000$0.0450
Unit Test Generation3,0005,000$0.0575
SQL Generation1,5001,000$0.0138
API Generation3,0006,000$0.0675
Code Refactoring4,0004,000$0.0500
Code Review5,0002,000$0.0325
PR Review8,0003,000$0.0500
Security Review6,0002,500$0.0400
Bug Detection5,0002,000$0.0325
Document Summarization10,0001,000$0.0350
PDF Summarization15,0001,500$0.0525
Meeting Notes8,0002,000$0.0400

Monthly cost projections

Projected monthly spend at different call volumes, using code generation as a representative workload (2,000 input / 4,000 output tokens per call).

Calls / monthWithout cachingWith caching (50% read discount)
1K$45$43.25
10K$450$432.5
100K$4,500$4,325
500K$22,500$21,625
1M$45,000$43,250

Caching estimate assumes 70% of input tokens are cache reads. Your actual cache hit rate will depend on prompt structure and reuse.

Mid-tier alternatives to GPT 4o

Same-tier models from other providers, sorted by input price.

ModelProviderInput $/MTokOutput $/MTokInput savings
Gemini 2.5 ProGemini$1.25$10.0050% cheaper
Grok 4.3Grok$1.25$2.5050% cheaper
Grok 4.20 Non-ReasoningGrok$1.25$2.5050% cheaper
Grok 4.20 ReasoningGrok$1.25$2.5050% cheaper
Grok 4.20 Multi-AgentGrok$1.25$2.5050% cheaper
Claude Sonnet 5Claude$2.00$10.0020% cheaper
Gemini 3.1 ProGemini$2.00$12.0020% cheaper
Gemini 3.1 Pro Custom ToolsGemini$2.00$12.0020% cheaper
Claude Sonnet 4.6Claude$3.00$15.0020% more

Other GPT models

Track your GPT 4o spend

See exactly how much you spend on GPT 4o and where you can save. Two-line wrapper install, no API keys shared.