GPT/Fast / Budget

GPT 5.6 Luna

GPT 5.6 Luna is GPT's fastest and most affordable model, priced at $0.20/MTok input and $1.20/MTok output with a 1,050K context window.

Input $/MTok

$0.20

Output $/MTok

$1.20

Context window

1,050K

Cache read discount

90% off

GPT 5.6 Luna pricing breakdown

RateInput $/MTokOutput $/MTok
Standard$0.20$1.20
Cached read (90% off input)$0.02-
Cached write (25% surcharge)$0.25-
Batch (50% off)$0.10$0.60
Long context (≥272.001K tokens)$0.40$1.80

All rates in USD per million tokens. Prices verified against provider documentation.

GPT 5.6 Luna task benchmarks

Confidence scores (0-100) and cross-provider rank for each task type. Rank 1 is best across all 20 models.

TaskConfidenceRank
Code Generation70/100#6
Code Review68/100#6
Summarization84/100#4
Q&A85/100#4
Extraction89/100#2
Reasoning63/100#6
Classification91/100#1
Creative Writing58/100#7

GPT 5.6 Luna cost by use case

Estimated cost per API call using default token counts for each workload.

Use caseInput tokensOutput tokensCost / call
Code Generation2,0004,000$0.005200
Unit Test Generation3,0005,000$0.006600
SQL Generation1,5001,000$0.001500
API Generation3,0006,000$0.007800
Code Refactoring4,0004,000$0.005600
Code Review5,0002,000$0.003400
PR Review8,0003,000$0.005200
Security Review6,0002,500$0.004200
Bug Detection5,0002,000$0.003400
Document Summarization10,0001,000$0.003200
PDF Summarization15,0001,500$0.004800
Meeting Notes8,0002,000$0.004000

Monthly cost projections

Projected monthly spend at different call volumes, using code generation as a representative workload (2,000 input / 4,000 output tokens per call).

Calls / monthWithout cachingWith caching (90% read discount)
1K$5.2$4.95
10K$52$49.48
100K$520$494.8
500K$2,600$2,474
1M$5,200$4,948

Caching estimate assumes 70% of input tokens are cache reads. Your actual cache hit rate will depend on prompt structure and reuse.

Fast / Budget alternatives to GPT 5.6 Luna

Same-tier models from other providers, sorted by input price.

ModelProviderInput $/MTokOutput $/MTokInput savings
Gemini 3.5 Flash-LiteGemini$0.30$2.5050% more
Gemini 3 FlashGemini$0.50$3.00150% more
Claude Haiku 4.5Claude$1.00$5.00400% more
Grok Build 0.1Grok$1.00$2.00400% more
Gemini 3.5 FlashGemini$1.50$9.00650% more

GPT 5.6 Luna pricing history

Input decreased

$1.00 to $0.20/MTok

Aug 13, 2026

Output decreased

$6.00 to $1.20/MTok

Aug 13, 2026

View full GPT 5.6 Luna pricing history

Track your GPT 5.6 Luna spend

See exactly how much you spend on GPT 5.6 Luna and where you can save. Two-line wrapper install, no API keys shared.