Gemini 3.5 Flash-Lite is Gemini's fastest and most affordable model, priced at $0.30/MTok input and $2.50/MTok output with a 1,000K context window.
Input $/MTok
$0.30
Output $/MTok
$2.50
Context window
1,000K
Cache read discount
90% off
| Rate | Input $/MTok | Output $/MTok |
|---|---|---|
| Standard | $0.30 | $2.50 |
| Cached read (90% off input) | $0.03 | - |
| Cached write (no surcharge) | $0.30 | - |
| Batch (50% off) | $0.15 | $1.25 |
All rates in USD per million tokens. Prices verified against provider documentation.
Confidence scores (0-100) and cross-provider rank for each task type. Rank 1 is best across all 20 models.
| Task | Confidence | Rank |
|---|---|---|
| Code Generation | 55/100 | #8 |
| Code Review | 52/100 | #8 |
| Summarization | 79/100 | #6 |
| Q&A | 78/100 | #6 |
| Extraction | 83/100 | #5 |
| Reasoning | 50/100 | #8 |
| Classification | 85/100 | #4 |
| Creative Writing | 45/100 | #9 |
Estimated cost per API call using default token counts for each workload.
| Use case | Input tokens | Output tokens | Cost / call |
|---|---|---|---|
| Code Generation | 2,000 | 4,000 | $0.0106 |
| Unit Test Generation | 3,000 | 5,000 | $0.0134 |
| SQL Generation | 1,500 | 1,000 | $0.002950 |
| API Generation | 3,000 | 6,000 | $0.0159 |
| Code Refactoring | 4,000 | 4,000 | $0.0112 |
| Code Review | 5,000 | 2,000 | $0.006500 |
| PR Review | 8,000 | 3,000 | $0.009900 |
| Security Review | 6,000 | 2,500 | $0.008050 |
| Bug Detection | 5,000 | 2,000 | $0.006500 |
| Document Summarization | 10,000 | 1,000 | $0.005500 |
| PDF Summarization | 15,000 | 1,500 | $0.008250 |
| Meeting Notes | 8,000 | 2,000 | $0.007400 |
Projected monthly spend at different call volumes, using code generation as a representative workload (2,000 input / 4,000 output tokens per call).
| Calls / month | Without caching | With caching (90% read discount) |
|---|---|---|
| 1K | $10.6 | $10.22 |
| 10K | $106 | $102.22 |
| 100K | $1,060 | $1,022.2 |
| 500K | $5,300 | $5,111 |
| 1M | $10,600 | $10,222 |
Caching estimate assumes 70% of input tokens are cache reads. Your actual cache hit rate will depend on prompt structure and reuse.
Same-tier models from other providers, sorted by input price.
| Model | Provider | Input $/MTok | Output $/MTok | Input savings |
|---|---|---|---|---|
| GPT 5.6 Luna | GPT | $0.20 | $1.20 | 33% cheaper |
| Claude Haiku 4.5 | Claude | $1.00 | $5.00 | 233% more |
| Grok Build 0.1 | Grok | $1.00 | $2.00 | 233% more |
See exactly how much you spend on Gemini 3.5 Flash-Lite and where you can save. Two-line wrapper install, no API keys shared.