June 10, 2026
On June 10, 2026, Google reduced Gemini 3.5 Flash output pricing from $0.75 to $0.60 per million tokens. That is a 20% cut on output, while input pricing remains unchanged at $1.50/MTok.
The current registry shows Gemini 3.5 Flash output at $9.00/MTok after a subsequent pricing restructure. This article covers the June 2026 change as recorded in the pricing changelog at that time.
Gemini 3.5 Flash ranks first for classification at 92/100 confidence and extraction at 90/100 confidence. It handles high-volume structured tasks at a fraction of what flagship models charge per call.
For a classification pipeline running 100,000 calls/day at 500 input and 200 output tokens:
Per-call savings look small, but at 100,000 calls per day the output cut saves ~$90/month.
Google now offers four models priced under $2/MTok input across the fast tier:
For workloads where Flash-tier confidence scores meet the bar, Google models are consistently cheapest.
If you run classification or extraction workloads on a flagship model, test Gemini 3.5 Flash instead. At 92/100 confidence for classification, Flash competes with models that cost 20x more per call.
Run a free cost comparison at /teardown, or start tracking your real spend with a free dashboard.
GPT-5.5 input cost dropped from $7.50 to $6.00 per million tokens in May 2026. What it means for teams running flagship workloads.
Anthropic raised Claude Opus 4.8 input pricing from $4.50 to $5.00 per million tokens on May 28, 2026. Here is what it means for flagship users.
Anthropic launched Claude Fable 5 on June 9, 2026 at $10.00 input and $50.00 output per million tokens. It is the most expensive model in the registry.
Calculators estimate. The dashboard shows what you actually spend and where you can save.