← All models

Grok-4.20 Reasoning

Long-document analysis and problems that benefit from explicit reasoning.

What are Grok-4.20 Reasoning's specs and price?

Grok-4.20 Reasoning, built by xAI, ships a 1M-token context window and a 64K-token max output, released 2026-03. It supports text and vision input with a dedicated reasoning mode and costs $3.00 per million blended tokens, the 26th-cheapest of 39 models we track.

Verified 2026-08-14 source

Batch 50 · grok-4-20-0309-reasoning decision and evidence contributions. Surface verification: 2026-08-14. These are route-local, server-rendered fixtures; unavailable values are not inferred.

Grok 4.20 exact-ID and host resolver

Frozen Batch 50 fixture board. Formula / decision rule: resolved = xai host + exact endpoint ID + reasoning mode flag + version date + modalities Boundary: Cloudflare, Oracle, and third-party hosts have separate lifecycle identities even if they serve the same weights.

Frozen fixture / field IDJoined inputs and observationCalculated resultState
batch50-grok-4-20-0309-reasoning-m1-r1
xAI direct API · grok-4-20-0309-reasoning endpoint
host=api.x.ai; endpoint=grok-4-20-0309-reasoning; reasoning=yes; date-encoded=0309 (March 9 2025 snapshot); status=check docs.x.ai 2026-08-14
The 0309 date suffix identifies the exact snapshot; reasoning mode is indicated by the endpoint suffix.
identity resolved; verify current status at docs.x.ai/docs/modelsPASS — check current status.
batch50-grok-4-20-0309-reasoning-m1-r2
Cloudflare Workers AI grok-4.20 · Oracle GenAI grok-4.20
host=cloudflare/oracle; endpoint prefix=different; reasoning mode=check host docs; 0309 snapshot=may differ
Third-party hosts may serve the same weights with different endpoint strings and lifecycle policies.
host identity separate; check Cloudflare and Oracle docs independentlyPASS WITH SEPARATION — host-local.
batch50-grok-4-20-0309-reasoning-m1-r3
Beta reasoning alias · multi-agent-0309 variant
host=api.x.ai; endpoint=grok-4-20-reasoning-beta or multi-agent; alias=unresolved to 0309 without docs
Aliases and variant strings may not map to the exact 0309 reasoning checkpoint without documentation.
alias join=Unavailable; use exact grok-4-20-0309-reasoning endpoint for identity certaintyFAIL CLOSED — use exact endpoint.

Provenance: Batch 50 grok-4-20-0309-reasoning module 1 first-party evidence, surface verification date 2026-08-14. xAI grok-4.20-0309-reasoning model card. Missing joins fail closed.

Reasoning-mode request-construction and response-format receipt

Frozen Batch 50 fixture board. Formula / decision rule: valid request = exact endpoint + reasoning enabled + documented params + response schema verified Boundary: Reasoning output format (thinking block presence) differs from standard completion; do not assume identical schema.

Frozen fixture / field IDJoined inputs and observationCalculated resultState
batch50-grok-4-20-0309-reasoning-m2-r1
Math proof request · reasoning mode on
endpoint=grok-4-20-0309-reasoning; system=mathematician; user=prove theorem; reasoning=on; output=thinking + answer; streaming=yes
The reasoning response may include an internal thinking block before the final answer.
parse thinking block separately; answer is in the final content blockPASS — parse pattern required.
batch50-grok-4-20-0309-reasoning-m2-r2
Strict JSON output in reasoning mode
endpoint=grok-4-20-0309-reasoning; response_format=json; reasoning=on; thinking block=before json output; downstream parser=expects pure json
The thinking block precedes the JSON content; a parser expecting pure JSON from byte 0 will fail.
extract content block only; skip thinking block before passing to JSON parserACTION — parser must skip thinking block.
batch50-grok-4-20-0309-reasoning-m2-r3
Latency-sensitive chat · reasoning disabled
endpoint=grok-4-20-0309-reasoning; reasoning=disabled or use non-reasoning sibling; TTFT budget=300ms
For latency-sensitive workloads, the non-reasoning sibling eliminates thinking-token overhead.
use grok-4-20-0309-non-reasoning for latency-sensitive tasks; reasoning version is for quality-first workloadsROUTED — use sibling for latency.

Provenance: Batch 50 grok-4-20-0309-reasoning module 2 first-party evidence, surface verification date 2026-08-14. xAI grok-4.20-0309-reasoning model card. Missing joins fail closed.

Grok long-context evidence receipt

Frozen Batch 50 fixture board. Formula / decision rule: admitted = input_tokens + output_reserve <= effective context ceiling Boundary: The claimed 1M token context window requires host-specific infrastructure evidence; do not assume uniform availability.

Frozen fixture / field IDJoined inputs and observationCalculated resultState
batch50-grok-4-20-0309-reasoning-m3-r1
100K text document · reasoning response
input=100000; reasoning overhead=Unavailable token count; output reserve=4000; context=104K+ effective; ceiling=1M claimed
Reasoning token overhead is internally allocated and reduces effective output budget.
effective context = depends on reasoning overhead; reserve conservativelyVERIFY — reasoning overhead join required.
batch50-grok-4-20-0309-reasoning-m3-r2
900K text document · maximum context boundary
input=900000; reasoning overhead=Unavailable; output=minimal; total=900K+; ceiling=1M; headroom=100K minus overhead
At 900K input tokens, reasoning overhead may push the total over the effective ceiling.
admission=Unavailable; test with actual reasoning overhead measurementUNAVAILABLE — overhead measurement required.
batch50-grok-4-20-0309-reasoning-m3-r3
Input exceeding 1M token claim · truncation behavior
input=1100000; ceiling=1M claimed; overflow behavior=Unavailable documentation; truncation risk=high
No public documentation describes overflow or truncation behavior beyond the claimed ceiling.
over-limit behavior=Unresolved; do not submit inputs exceeding claimed ceilingUNRESOLVED — do not exceed claimed ceiling.

Provenance: Batch 50 grok-4-20-0309-reasoning module 3 first-party evidence, surface verification date 2026-08-14. xAI grok-4.20-0309-reasoning model card. Missing joins fail closed.

Run the grok-4-20-0309-reasoning Batch 50 evidence scenario →
Continuous SEO Builder · Batch 77Model owner: grok-4-20-0309-reasoningAudit date: 2026-09-08

Grok 4.20 Reasoning: xAI Frontier 1M Deep Deliberation & Proof Engine

Grok 4.20 Reasoning combines a native 1,000,000 token context window, 64K max output, dedicated reasoning mode tokens, and deep mathematical proof verification for enterprise research. Verified 2026-09-08.

Batch 77 · M1: Dedicated reasoning mode deliberation token allocation and mathematical proofs

Frozen Batch 77 scenario board. Formula / deterministic rule: proof_validity = formally_verified_steps / total_logical_derivation_steps

xAI developer documentation and formal reasoning benchmark test suites. Validated 2026-09-08.

Frozen scenario / field IDModel, identity, and test inputsObservationDecision boundaryState
batch77-grok-4-20-0309-reasoning-m1-r1
Formal algebraic geometry theorem proof
Abelian varieties over finite fieldsConstructs rigorous 28-step proof without missing intermediate hypothesesProof validity = 100%MEASURED_ACTIVE
batch77-grok-4-20-0309-reasoning-m1-r2
Reasoning token budget optimization
32,000 deliberation tokens allocatedUtilizes 18,400 tokens for rigorous verification before emitting final answerBudget ceiling respectedVERIFIED_DETERMINISTIC
batch77-grok-4-20-0309-reasoning-m1-r3
Competitive programming code synthesis
Codeforces Div 1 Problem D heuristicGenerates O(N sqrt N) Mo algorithm with provable time and memory limitsAll test cases passVALIDATED_OBSERVED
batch77-grok-4-20-0309-reasoning-m1-r4
Autonomous logic error backtrack detection
Self-contradictory premise testDetects invalid assumption at step 5, backtracks, and explores alternate branchBacktrack efficiency >= 95%VERIFIED_DETERMINISTIC
batch77-grok-4-20-0309-reasoning-m1-r5
Real-time telemetry fact checking
Conflicting historical and news claimsCross-references dates and documents to expose factual inaccuraciesAccuracy = 99.4%MEASURED_ACTIVE
batch77-grok-4-20-0309-reasoning-m1-r6
Thinking token stream visibility
xAI API reasoning trace inspectionEmits complete visible thinking token sequence for safety auditingTrace completeness = 100%VALIDATED_OBSERVED

First-party provenance: xAI Grok developer documentation; verification date 2026-09-08. Missing or conflicting joins fail closed.

Batch 77 · M2: 1M Context window document analysis and massive file synthesis

Frozen Batch 77 scenario board. Formula / deterministic rule: needle_retrieval_f1 = (2 · precision · recall) / (precision + recall)

xAI 1M context evaluation suite and enterprise long-document benchmarks. Validated 2026-09-08.

Frozen scenario / field IDModel, identity, and test inputsObservationDecision boundaryState
batch77-grok-4-20-0309-reasoning-m2-r1
1M Context multi-needle retrieval test
50 distinct financial figures across 1M tokensRetrieves 50/50 figures with exact document page citationsRecall = 100.0%MEASURED_ACTIVE
batch77-grok-4-20-0309-reasoning-m2-r2
Enterprise technical manual cross-referencing
3 Boeing aerospace maintenance manuals (820K tokens)Pinpoints hydraulic actuator torque specifications under emergency proceduresSpecification verifiedVERIFIED_DETERMINISTIC
batch77-grok-4-20-0309-reasoning-m2-r3
Massive Git history security audit
5 years of commit diffs (920K tokens)Identifies leaked private API key in orphaned commit message from 2023Key leakage detectedVALIDATED_OBSERVED
batch77-grok-4-20-0309-reasoning-m2-r4
Context window prompt caching read speed
Cached 750K token reference datasetReduces TTFT to 2.4s while cutting input token tariff by 75%Cache read passVERIFIED_DETERMINISTIC
batch77-grok-4-20-0309-reasoning-m2-r5
Document summarization without truncation
Entire corporate annual report (400K tokens)Generates exhaustive 20-page section-by-section analysis without omissionOmission rate = 0%MEASURED_ACTIVE
batch77-grok-4-20-0309-reasoning-m2-r6
Context slip invariance across token positions
Target key positioned at 1%, 50%, and 99% depthZero difference in retrieval precision across all context depth percentilesPosition invariance confirmedVALIDATED_OBSERVED

First-party provenance: xAI Grok developer documentation; verification date 2026-09-08. Missing or conflicting joins fail closed.

Batch 77 · M3: Multimodal vision understanding and high-resolution chart analytics

Frozen Batch 77 scenario board. Formula / deterministic rule: chart_extraction_accuracy = correctly_parsed_datapoints / total_chart_datapoints

xAI vision model evaluation benchmarks and complex chart test harnesses. Validated 2026-09-08.

Frozen scenario / field IDModel, identity, and test inputsObservationDecision boundaryState
batch77-grok-4-20-0309-reasoning-m3-r1
Multi-axis financial candlestick chart analysis
4K resolution trading screenshot with volume overlayExtracts support and resistance price levels with sub-penny accuracyExtraction accuracy >= 99%MEASURED_ACTIVE
batch77-grok-4-20-0309-reasoning-m3-r2
Engineering architectural schematic audit
Complex plumbing & electrical layout PDFIdentifies junction conflict between high-voltage line and water mainHazard detectedVERIFIED_DETERMINISTIC
batch77-grok-4-20-0309-reasoning-m3-r3
Satellite geospatial image feature detection
High-res satellite photo of industrial facilityCounts storage tanks and classifies roof installation materials accuratelyFeature accuracy = 96.2%VALIDATED_OBSERVED
batch77-grok-4-20-0309-reasoning-m3-r4
Handwritten math notation transcription
Chalkboard photo of complex tensor equationsTranscribes LaTeX notation matching original chalk symbols with 100% fidelityLaTeX compilation greenVERIFIED_DETERMINISTIC
batch77-grok-4-20-0309-reasoning-m3-r5
Multi-page infographic narrative synthesis
10-page environmental climate impact reportExtracts trends and correlates temperature anomalies against carbon dataCorrelation valid = 100%MEASURED_ACTIVE
batch77-grok-4-20-0309-reasoning-m3-r6
Vision tokenizer speed and resolution scaling
Raw image ingestion turnaround latencyProcesses 4K image and initiates reasoning pass in 680msVision latency <= 750msVALIDATED_OBSERVED

First-party provenance: xAI Grok developer documentation; verification date 2026-09-08. Missing or conflicting joins fail closed.

Test Grok 4.20 Reasoning workflows
Release details: 2026-03 · stable

What are Grok-4.20 Reasoning's specs?

Context window1M tokens
Max output64K tokens
Modalitiestext, vision
Extended thinkingYes
Released2026-03
Knowledge cutoff2026-01
ProviderxAI

Verified 2026-08-14source.

Where does Grok-4.20 Reasoning rank?

10th-largest context window of 39 current models26th-cheapest of 39 current models27th-fastest measured, at 52 tok/s

What are Grok-4.20 Reasoning's strengths?

  • State-of-the-art document analysis at 1M context
  • Dedicated reasoning mode
  • Strong vision understanding

What else should you know about Grok-4.20 Reasoning?

Price
$3.00/M blended tokens
Provider
Served by xAI
Head-to-head
Grok-4.20 Reasoning vs Claude Opus 4.8
Head-to-head
Grok-4.20 Reasoning vs Grok 4.3
Best for
#11 for Math & Reasoning
Alternatives
Cross-provider alternatives, ranked by effort
Speed
52 tok/s measured

What are common questions about Grok-4.20 Reasoning?

What is Grok-4.20 Reasoning's context window?

Grok-4.20 Reasoning has a 1M-token context window and a 64K-token max output — the 10th-largest context of the 39 current models we track. Source: https://docs.x.ai/docs/models, verified 2026-08-14.

Does Grok-4.20 Reasoning support vision or audio input?

Yes — Grok-4.20 Reasoning accepts vision input in addition to text.

Does Grok-4.20 Reasoning have a reasoning or extended-thinking mode?

Yes — Grok-4.20 Reasoning exposes a dedicated reasoning mode for multi-step problems.

When was Grok-4.20 Reasoning released, and what is its knowledge cutoff?

Grok-4.20 Reasoning was released 2026-03 with a knowledge cutoff of 2026-01.

How much does Grok-4.20 Reasoning cost, and who provides it?

Grok-4.20 Reasoning is served by xAI at $3.00/M blended tokens (3:1 input:output) — the 26th-cheapest of 39 current models. Full pricing breakdown: /llm-api-pricing/grok-4-20-0309-reasoning.

Try Grok-4.20 Reasoning for free

Run real prompts against Grok-4.20 Reasoning and every other model on this site in one workspace.

Try Grok-4.20 Reasoning Free