Gemini/Mid-tier

Gemini 3.1 Pro Custom Tools

Gemini 3.1 Pro Custom Tools is Gemini's balanced performance model, priced at $2.00/MTok input and $12.00/MTok output with a 1,048.576K context window. Backfilled on 2026-09-23. Prices confirmed by the provider pricing page and the LiteLLM cross-check. Release date not published in a machine-readable source. Not yet benchmarked, so never recommended.

Input $/MTok

$2.00

Output $/MTok

$12.00

Context window

1,048.576K

Cache read discount

90% off

Gemini 3.1 Pro Custom Tools pricing breakdown

RateInput $/MTokOutput $/MTok
Standard$2.00$12.00
Cached read (90% off input)$0.20-
Cached write (no surcharge)$2.00-
Batch (50% off)$1.00$6.00
Long context (≥200.001K tokens)$4.00$18.00

All rates in USD per million tokens. Prices verified against provider documentation.

Gemini 3.1 Pro Custom Tools task benchmarks

Confidence scores (0-100) and cross-provider rank for each task type. Rank 1 is best across all 72 models.

Not benchmarked. The third-party benchmark source used for confidence scores does not cover Gemini models yet. Scores will appear here once independent benchmarks are published.

Gemini 3.1 Pro Custom Tools cost by use case

Estimated cost per API call using default token counts for each workload.

Use caseInput tokensOutput tokensCost / call
Code Generation2,0004,000$0.0520
Unit Test Generation3,0005,000$0.0660
SQL Generation1,5001,000$0.0150
API Generation3,0006,000$0.0780
Code Refactoring4,0004,000$0.0560
Code Review5,0002,000$0.0340
PR Review8,0003,000$0.0520
Security Review6,0002,500$0.0420
Bug Detection5,0002,000$0.0340
Document Summarization10,0001,000$0.0320
PDF Summarization15,0001,500$0.0480
Meeting Notes8,0002,000$0.0400

Monthly cost projections

Projected monthly spend at different call volumes, using code generation as a representative workload (2,000 input / 4,000 output tokens per call).

Calls / monthWithout cachingWith caching (90% read discount)
1K$52$49.48
10K$520$494.8
100K$5,200$4,948
500K$26,000$24,740
1M$52,000$49,480

Caching estimate assumes 70% of input tokens are cache reads. Your actual cache hit rate will depend on prompt structure and reuse.

Mid-tier alternatives to Gemini 3.1 Pro Custom Tools

Same-tier models from other providers, sorted by input price.

ModelProviderInput $/MTokOutput $/MTokInput savings
GPT 5GPT$1.25$10.0038% cheaper
GPT 5.1GPT$1.25$10.0038% cheaper
Grok 4.3Grok$1.25$2.5038% cheaper
Grok 4.20 Non-ReasoningGrok$1.25$2.5038% cheaper
Grok 4.20 ReasoningGrok$1.25$2.5038% cheaper
Grok 4.20 Multi-AgentGrok$1.25$2.5038% cheaper
GPT 5.2GPT$1.75$14.0013% cheaper
Claude Sonnet 5Claude$2.00$10.00Same price
GPT 5.6 TerraGPT$2.00$12.00Same price
GPT 4.1GPT$2.00$8.00Same price
o3GPT$2.00$8.00Same price
GPT 5.4GPT$2.50$15.0025% more
GPT 4oGPT$2.50$10.0025% more
Claude Sonnet 4.6Claude$3.00$15.0050% more

Track your Gemini 3.1 Pro Custom Tools spend

See exactly how much you spend on Gemini 3.1 Pro Custom Tools and where you can save. Two-line wrapper install, no API keys shared.