Gemini/Fast / Budget

Gemini Omni 1.1 Flash

Gemini Omni 1.1 Flash is Gemini's fastest and most affordable model, priced at $1.50/MTok input and $9.00/MTok output with a 1,048.576K context window. Backfilled on 2026-09-23. Prices confirmed by the provider pricing page and the LiteLLM cross-check. Release date not published in a machine-readable source. Not yet benchmarked, so never recommended.

Input $/MTok

$1.50

Output $/MTok

$9.00

Context window

1,048.576K

Cache read discount

0% off

Gemini Omni 1.1 Flash pricing breakdown

RateInput $/MTokOutput $/MTok
Standard$1.50$9.00
Cached read (0% off input)$1.50-
Cached write (no surcharge)$1.50-
Batch (50% off)$0.75$4.50

All rates in USD per million tokens. Prices verified against provider documentation.

Gemini Omni 1.1 Flash task benchmarks

Confidence scores (0-100) and cross-provider rank for each task type. Rank 1 is best across all 72 models.

Not benchmarked. The third-party benchmark source used for confidence scores does not cover Gemini models yet. Scores will appear here once independent benchmarks are published.

Gemini Omni 1.1 Flash cost by use case

Estimated cost per API call using default token counts for each workload.

Use caseInput tokensOutput tokensCost / call
Code Generation2,0004,000$0.0390
Unit Test Generation3,0005,000$0.0495
SQL Generation1,5001,000$0.0113
API Generation3,0006,000$0.0585
Code Refactoring4,0004,000$0.0420
Code Review5,0002,000$0.0255
PR Review8,0003,000$0.0390
Security Review6,0002,500$0.0315
Bug Detection5,0002,000$0.0255
Document Summarization10,0001,000$0.0240
PDF Summarization15,0001,500$0.0360
Meeting Notes8,0002,000$0.0300

Monthly cost projections

Projected monthly spend at different call volumes, using code generation as a representative workload (2,000 input / 4,000 output tokens per call).

Calls / monthWithout cachingWith caching (0% read discount)
1K$39$39
10K$390$390
100K$3,900$3,900
500K$19,500$19,500
1M$39,000$39,000

Caching estimate assumes 70% of input tokens are cache reads. Your actual cache hit rate will depend on prompt structure and reuse.

Fast / Budget alternatives to Gemini Omni 1.1 Flash

Same-tier models from other providers, sorted by input price.

ModelProviderInput $/MTokOutput $/MTokInput savings
GPT 5 NanoGPT$0.05$0.4097% cheaper
GPT 4.1 NanoGPT$0.10$0.4093% cheaper
GPT 6 LunaGPT$0.10$0.5093% cheaper
GPT 4o MiniGPT$0.15$0.6090% cheaper
GPT 5.6 LunaGPT$0.20$1.2087% cheaper
GPT 5.4 NanoGPT$0.20$1.2587% cheaper
GPT 5 MiniGPT$0.25$2.0083% cheaper
GPT 4.1 MiniGPT$0.40$1.6073% cheaper
GPT 3.5 TurboGPT$0.50$1.5067% cheaper
GPT 3.5 Turbo (0125)GPT$0.50$1.5067% cheaper
GPT 5.4 MiniGPT$0.75$4.5050% cheaper
Claude Haiku 4.5Claude$1.00$5.0033% cheaper
Grok Build 0.1Grok$1.00$2.0033% cheaper
o3 MiniGPT$1.10$4.4027% cheaper
o4 MiniGPT$1.10$4.4027% cheaper

Track your Gemini Omni 1.1 Flash spend

See exactly how much you spend on Gemini Omni 1.1 Flash and where you can save. Two-line wrapper install, no API keys shared.