July 10, 2026
On July 9, 2026, OpenAI released the GPT 5.6 family with three tiers:
All three models ship with 1,000,000 token context windows. That matches Gemini and exceeds GPT 5.5 at 128K.
Flagship tier:
Mid tier:
Fast tier:
Terra scores 92/100 on summarization and 92/100 on QA, ranking first in both tasks. At $2.50/MTok input, it undercuts Claude Sonnet 5 by $0.50 on input pricing.
For a RAG pipeline at 6,000 input and 1,500 output tokens per call:
Gemini 3.1 Pro remains the cheapest mid-tier option for most workloads.
GPT 5.6 Luna matches Claude Haiku 4.5 at $1.00/MTok input. Luna scores 91/100 on classification versus Haiku at 90/100. Luna costs $1.00 more on output per MTok.
The tradeoff: Luna is better at classification, Haiku is cheaper on output-heavy workloads.
Compare the full 5.6 family against your current models at /teardown, or start tracking with a free dashboard.
GPT-5.5 input cost dropped from $7.50 to $6.00 per million tokens in May 2026. What it means for teams running flagship workloads.
Gemini 3.5 Flash output pricing fell from $0.75 to $0.60 per million tokens in June 2026, widening Google's cost advantage in the fast tier.
Anthropic raised Claude Opus 4.8 input pricing from $4.50 to $5.00 per million tokens on May 28, 2026. Here is what it means for flagship users.
Calculators estimate. The dashboard shows what you actually spend and where you can save.