Gemini/Fast / Budget

Gemini 2.5 Flash-Lite

Gemini 2.5 Flash-Lite is Gemini's fastest and most affordable model, priced at $0.10/MTok input and $0.40/MTok output with a 1,048.576K context window. Backfilled on 2026-09-23. Prices confirmed by the provider pricing page and the LiteLLM cross-check. Release date not published in a machine-readable source. Not yet benchmarked, so never recommended.

Input $/MTok

$0.10

Output $/MTok

$0.40

Context window

1,048.576K

Cache read discount

90% off

Gemini 2.5 Flash-Lite pricing breakdown

RateInput $/MTokOutput $/MTok
Standard$0.10$0.40
Cached read (90% off input)$0.01-
Cached write (no surcharge)$0.10-
Batch (50% off)$0.05$0.20

All rates in USD per million tokens. Prices verified against provider documentation.

Gemini 2.5 Flash-Lite task benchmarks

Confidence scores (0-100) and cross-provider rank for each task type. Rank 1 is best across all 72 models.

Not benchmarked. The third-party benchmark source used for confidence scores does not cover Gemini models yet. Scores will appear here once independent benchmarks are published.

Gemini 2.5 Flash-Lite cost by use case

Estimated cost per API call using default token counts for each workload.

Use caseInput tokensOutput tokensCost / call
Code Generation2,0004,000$0.001800
Unit Test Generation3,0005,000$0.002300
SQL Generation1,5001,000$0.000550
API Generation3,0006,000$0.002700
Code Refactoring4,0004,000$0.002000
Code Review5,0002,000$0.001300
PR Review8,0003,000$0.002000
Security Review6,0002,500$0.001600
Bug Detection5,0002,000$0.001300
Document Summarization10,0001,000$0.001400
PDF Summarization15,0001,500$0.002100
Meeting Notes8,0002,000$0.001600

Monthly cost projections

Projected monthly spend at different call volumes, using code generation as a representative workload (2,000 input / 4,000 output tokens per call).

Calls / monthWithout cachingWith caching (90% read discount)
1K$1.8$1.67
10K$18$16.74
100K$180$167.4
500K$900$837
1M$1,800$1,674

Caching estimate assumes 70% of input tokens are cache reads. Your actual cache hit rate will depend on prompt structure and reuse.

Fast / Budget alternatives to Gemini 2.5 Flash-Lite

Same-tier models from other providers, sorted by input price.

ModelProviderInput $/MTokOutput $/MTokInput savings
GPT 5 NanoGPT$0.05$0.4050% cheaper
GPT 4.1 NanoGPT$0.10$0.40Same price
GPT 6 LunaGPT$0.10$0.50Same price
GPT 4o MiniGPT$0.15$0.6050% more
GPT 5.6 LunaGPT$0.20$1.20100% more
GPT 5.4 NanoGPT$0.20$1.25100% more
GPT 5 MiniGPT$0.25$2.00150% more
GPT 4.1 MiniGPT$0.40$1.60300% more
GPT 3.5 TurboGPT$0.50$1.50400% more
GPT 3.5 Turbo (0125)GPT$0.50$1.50400% more
GPT 5.4 MiniGPT$0.75$4.50650% more
Claude Haiku 4.5Claude$1.00$5.00900% more
Grok Build 0.1Grok$1.00$2.00900% more
o3 MiniGPT$1.10$4.401000% more
o4 MiniGPT$1.10$4.401000% more

Track your Gemini 2.5 Flash-Lite spend

See exactly how much you spend on Gemini 2.5 Flash-Lite and where you can save. Two-line wrapper install, no API keys shared.