GPT 5.6 Sol is GPT's most capable model, priced at $4.00/MTok input and $20.00/MTok output with a 1,050K context window.
Input $/MTok
$4.00
Output $/MTok
$20.00
Context window
1,050K
Cache read discount
90% off
| Rate | Input $/MTok | Output $/MTok |
|---|---|---|
| Standard | $4.00 | $20.00 |
| Cached read (90% off input) | $0.40 | - |
| Cached write (25% surcharge) | $5.00 | - |
| Batch (50% off) | $2.00 | $10.00 |
| Long context (≥272.001K tokens) | $8.00 | $30.00 |
All rates in USD per million tokens. Prices verified against provider documentation.
Confidence scores (0-100) and cross-provider rank for each task type. Rank 1 is best across all 20 models.
| Task | Confidence | Rank |
|---|---|---|
| Code Generation | 94/100 | #1 |
| Code Review | 92/100 | #2 |
| Summarization | 90/100 | #2 |
| Q&A | 91/100 | #1 |
| Extraction | 88/100 | #3 |
| Reasoning | 96/100 | #1 |
| Classification | 86/100 | #3 |
| Creative Writing | 92/100 | #2 |
Estimated cost per API call using default token counts for each workload.
| Use case | Input tokens | Output tokens | Cost / call |
|---|---|---|---|
| Code Generation | 2,000 | 4,000 | $0.0880 |
| Unit Test Generation | 3,000 | 5,000 | $0.1120 |
| SQL Generation | 1,500 | 1,000 | $0.0260 |
| API Generation | 3,000 | 6,000 | $0.1320 |
| Code Refactoring | 4,000 | 4,000 | $0.0960 |
| Code Review | 5,000 | 2,000 | $0.0600 |
| PR Review | 8,000 | 3,000 | $0.0920 |
| Security Review | 6,000 | 2,500 | $0.0740 |
| Bug Detection | 5,000 | 2,000 | $0.0600 |
| Document Summarization | 10,000 | 1,000 | $0.0600 |
| PDF Summarization | 15,000 | 1,500 | $0.0900 |
| Meeting Notes | 8,000 | 2,000 | $0.0720 |
Projected monthly spend at different call volumes, using code generation as a representative workload (2,000 input / 4,000 output tokens per call).
| Calls / month | Without caching | With caching (90% read discount) |
|---|---|---|
| 1K | $88 | $82.96 |
| 10K | $880 | $829.6 |
| 100K | $8,800 | $8,296 |
| 500K | $44,000 | $41,480 |
| 1M | $88,000 | $82,960 |
Caching estimate assumes 70% of input tokens are cache reads. Your actual cache hit rate will depend on prompt structure and reuse.
Same-tier models from other providers, sorted by input price.
| Model | Provider | Input $/MTok | Output $/MTok | Input savings |
|---|---|---|---|---|
| Grok 4.6 | Grok | $2.00 | $6.00 | 50% cheaper |
| Grok 4.5 | Grok | $2.00 | $6.00 | 50% cheaper |
| Claude Opus 4.8 | Claude | $5.00 | $25.00 | 25% more |
| Claude Opus 5 | Claude | $5.00 | $25.00 | 25% more |
| Claude Fable 5.1 | Claude | $10.00 | $50.00 | 150% more |
| Claude Fable 5 | Claude | $10.00 | $50.00 | 150% more |
Input decreased
$5.00 to $4.00/MTok
Output decreased
$30.00 to $20.00/MTok
Input launched
$5.00/MTok at launch
See exactly how much you spend on GPT 5.6 Sol and where you can save. Two-line wrapper install, no API keys shared.