Grok/Mid-tier

Grok 4.20 Non-Reasoning

Grok 4.20 Non-Reasoning is Grok's balanced performance model, priced at $1.25/MTok input and $2.50/MTok output with a 1,000K context window. Backfilled on 2026-09-23. Prices confirmed by the provider pricing page and the LiteLLM cross-check. Release date not published in a machine-readable source. Not yet benchmarked, so never recommended.

Input $/MTok

$1.25

Output $/MTok

$2.50

Context window

1,000K

Cache read discount

84% off

Grok 4.20 Non-Reasoning pricing breakdown

RateInput $/MTokOutput $/MTok
Standard$1.25$2.50
Cached read (84% off input)$0.20-
Cached write (no surcharge)$1.25-
Batch (50% off)$0.63$1.25
Long context (≥200K tokens)$2.50$5.00

All rates in USD per million tokens. Prices verified against provider documentation.

Grok 4.20 Non-Reasoning task benchmarks

Confidence scores (0-100) and cross-provider rank for each task type. Rank 1 is best across all 72 models.

Not benchmarked. The third-party benchmark source used for confidence scores does not cover Grok models yet. Scores will appear here once independent benchmarks are published.

Grok 4.20 Non-Reasoning cost by use case

Estimated cost per API call using default token counts for each workload.

Use caseInput tokensOutput tokensCost / call
Code Generation2,0004,000$0.0125
Unit Test Generation3,0005,000$0.0163
SQL Generation1,5001,000$0.004375
API Generation3,0006,000$0.0187
Code Refactoring4,0004,000$0.0150
Code Review5,0002,000$0.0112
PR Review8,0003,000$0.0175
Security Review6,0002,500$0.0138
Bug Detection5,0002,000$0.0112
Document Summarization10,0001,000$0.0150
PDF Summarization15,0001,500$0.0225
Meeting Notes8,0002,000$0.0150

Monthly cost projections

Projected monthly spend at different call volumes, using code generation as a representative workload (2,000 input / 4,000 output tokens per call).

Calls / monthWithout cachingWith caching (84% read discount)
1K$12.5$11.03
10K$125$110.3
100K$1,250$1,103
500K$6,250$5,515
1M$12,500$11,030

Caching estimate assumes 70% of input tokens are cache reads. Your actual cache hit rate will depend on prompt structure and reuse.

Mid-tier alternatives to Grok 4.20 Non-Reasoning

Same-tier models from other providers, sorted by input price.

ModelProviderInput $/MTokOutput $/MTokInput savings
GPT 5GPT$1.25$10.00Same price
GPT 5.1GPT$1.25$10.00Same price
Gemini 2.5 ProGemini$1.25$10.00Same price
GPT 5.2GPT$1.75$14.0040% more
Claude Sonnet 5Claude$2.00$10.0060% more
GPT 5.6 TerraGPT$2.00$12.0060% more
GPT 4.1GPT$2.00$8.0060% more
o3GPT$2.00$8.0060% more
Gemini 3.1 ProGemini$2.00$12.0060% more
Gemini 3.1 Pro Custom ToolsGemini$2.00$12.0060% more
GPT 5.4GPT$2.50$15.00100% more
GPT 4oGPT$2.50$10.00100% more
Claude Sonnet 4.6Claude$3.00$15.00140% more

Track your Grok 4.20 Non-Reasoning spend

See exactly how much you spend on Grok 4.20 Non-Reasoning and where you can save. Two-line wrapper install, no API keys shared.