← All models

Grok-4.20

Fast multimodal chat and everyday tasks that don’t need extended reasoning.

Grok-4.20 supersedes Grok-3, Grok-3 Mini.

What are Grok-4.20's specs and price?

Grok-4.20, built by xAI, ships a 1M-token context window and a 32K-token max output, released 2026-03. It supports text and vision input and costs $3.00 per million blended tokens, the 27th-cheapest of 39 models we track.

Verified 2026-09-01 source

Batch 51 · grok-4-20-0309-non-reasoning evidence contributions. Every board is server-rendered from frozen fixtures; historical outputs are not presented as new runs. Verification date: 2026-09-01.

Non-reasoning request-control validator

Frozen Batch 51 fixture board. Formula / decision rule: accept only when exact endpoint + documented control + mode join is true; otherwise reject or mark Unresolved Boundary: A reasoning sibling or unsupported control never becomes accepted through silent reinterpretation.

Frozen fixture / field IDIdentity keysDeterministic ruleOutput / bounded stateValidation
batch51-grok-4-20-0309-non-reasoning-m1-r1
ordinary response
provider=xAI; host=api.x.ai; requested ID=grok-4.20-0309-non-reasoning; mode=non-reasoning; run=2026-09-01exact ID and ordinary response control must join firstaccepted envelope; response-field expectation=ordinary text outputACCEPTED — exact mode join
batch51-grok-4-20-0309-non-reasoning-m1-r2
explicit reasoning-effort parameter
provider=xAI; host=api.x.ai; requested ID=grok-4.20-0309-non-reasoning; control=reasoning_effort; value=explicit; documentation=mode-specifica non-reasoning endpoint must reject a reasoning-only control rather than change modereject request; do not reinterpret as reasoning endpointREJECTED — mode mismatch
batch51-grok-4-20-0309-non-reasoning-m1-r3
reasoning-only alias
provider=xAI; host=api.x.ai; requested ID=grok-4.20-0309-reasoning; alias=reasoning-only; target owner=separatereasoning alias is not the non-reasoning endpointroute to exact sibling owner or reject; no fact transferREJECTED — sibling identity
batch51-grok-4-20-0309-non-reasoning-m1-r4
strict JSON schema
provider=xAI; host=api.x.ai; endpoint=non-reasoning; response_format=strict JSON; schema hash=Unavailableschema acceptance requires an exact documented response contract and joined schemaschema result=Unavailable; preserve request for verificationUNRESOLVED — schema join
batch51-grok-4-20-0309-non-reasoning-m1-r5
image-plus-tool call
provider=xAI; host=api.x.ai; endpoint=non-reasoning; modality=image; tool=declared; tool schema hash=Unavailableimage, tool, and schema controls must each be documented for this endpointadmission=Unavailable; do not infer combined supportFAIL CLOSED — combined join
batch51-grok-4-20-0309-non-reasoning-m1-r6
unsupported custom-control fixtures
provider=xAI; host=api.x.ai; endpoint=non-reasoning; controls=custom deliberation budget and hidden mode flagunknown controls are not defaultsreject unknown controls; retain the original request envelopeREJECTED — unsupported control

Provenance: Batch 51 grok-4-20-0309-non-reasoning module 1; audit verification 2026-09-01. xAI Grok model documentation. Missing or conflicting joins fail closed.

Fast-model context admission receipt

Frozen Batch 51 fixture board. Formula / decision rule: admit = known input accounting + requested output + tool/schema reserve <= documented context; unknown accounting => Unresolved Boundary: No cost or quota is calculated here, and undocumented image/tool accounting is never treated as zero.

Frozen fixture / field IDIdentity keysDeterministic ruleOutput / bounded stateValidation
batch51-grok-4-20-0309-non-reasoning-m2-r1
8K chat
provider=xAI; host=api.x.ai; endpoint=grok-4.20-0309-non-reasoning; text=8K; output reserve=2K; context=first-party jointext input + output reserve must fit the exact endpoint ceilingadmission=within documented text budget; reserve retainedADMITTED — text accounting joined
batch51-grok-4-20-0309-non-reasoning-m2-r2
190K retrieval
provider=xAI; host=api.x.ai; endpoint=exact dated ID; text retrieval=190K; tools=no; output reserve=4Kknown text estimate and output reserve are compared without tariff arithmeticadmission=eligible if exact host context join remains currentADMITTED WITH DATE GATE
batch51-grok-4-20-0309-non-reasoning-m2-r3
201K threshold crossing
provider=xAI; host=api.x.ai; endpoint=exact ID; text=201K; output reserve=4K; context threshold=host-qualifiedcrossing a documented threshold requires the exact host and snapshot joinadmission=Unresolved until threshold documentation is rejoinedUNRESOLVED — threshold join
batch51-grok-4-20-0309-non-reasoning-m2-r4
900K document
provider=xAI; host=api.x.ai; endpoint=exact ID; text=900K; output reserve=8K; context claim=1Mtext total plus reserve is checked against the dated claimadmission=bounded by dated context evidence; no throughput claimADMITTED WITH LIMIT
batch51-grok-4-20-0309-non-reasoning-m2-r5
image-plus-700K text
provider=xAI; host=api.x.ai; endpoint=exact ID; text=700K; image=1; image units=Unavailable; output reserve=4Kunknown non-text units cannot be set to zeroadmission=Unavailable; verify image accounting before submissionFAIL CLOSED — media accounting
batch51-grok-4-20-0309-non-reasoning-m2-r6
over-1M fixtures
provider=xAI; host=api.x.ai; endpoint=exact ID; text>1M; overflow behavior=Unavailableinputs beyond the documented ceiling are not clipped or retried by assumptiondo not submit; truncation/error state=UnavailableREJECTED — over documented ceiling

Provenance: Batch 51 grok-4-20-0309-non-reasoning module 2; audit verification 2026-09-01. xAI Grok model documentation. Missing or conflicting joins fail closed.

Tool-call replay and side-effect board

Frozen Batch 51 fixture board. Formula / decision rule: replay = completed + schema-valid + idempotency-safe + side-effect policy allows replay Boundary: Endpoint availability does not make a payment-like or duplicate write safe to replay.

Frozen fixture / field IDIdentity keysDeterministic ruleOutput / bounded stateValidation
batch51-grok-4-20-0309-non-reasoning-m3-r1
read-only search
provider=xAI; host=api.x.ai; exact ID; tool_call_id=search-001; side effect=read; idempotency=not requiredread-only completed calls may be replayed when schema remains validreplay=allowed after response/schema validationSAFE — read-only replay
batch51-grok-4-20-0309-non-reasoning-m3-r2
parallel retrieval
provider=xAI; host=api.x.ai; exact ID; tool_call_ids=ret-001/ret-002; side effect=read; stream=completeeach tool call needs its own completion and schema joinreplay independently only after both calls validateCONDITIONAL — per-call validation
batch51-grok-4-20-0309-non-reasoning-m3-r3
strict-schema lookup
provider=xAI; host=api.x.ai; exact ID; tool_call_id=lookup-003; schema= strict; validation=passstrict schema pass plus same tool identity is requiredreplay=allowed; preserve tool-call ID lineageSAFE — schema-valid
batch51-grok-4-20-0309-non-reasoning-m3-r4
payment-like write
provider=xAI; host=api.x.ai; exact ID; tool_call_id=pay-004; side effect=external write; idempotency key=missingexternal writes require an idempotency key and operator containmentblock replay; require human/operator confirmationBLOCKED — unsafe side effect
batch51-grok-4-20-0309-non-reasoning-m3-r5
partial streamed tool call
provider=xAI; host=api.x.ai; exact ID; tool_call_id=stream-005; stream completion=partial; arguments=partialpartial arguments are not a completed tool invocationdo not execute or replay; await/mark incompleteBLOCKED — incomplete stream
batch51-grok-4-20-0309-non-reasoning-m3-r6
duplicate retry fixtures
provider=xAI; host=api.x.ai; exact ID; tool_call_id=retry-006; prior completion=unknown; idempotency=duplicateunknown prior completion makes a retry unsafe for side effectscontain and reconcile before retry; no automatic duplicate callBLOCKED — duplicate risk

Provenance: Batch 51 grok-4-20-0309-non-reasoning module 3; audit verification 2026-09-01. xAI Grok model documentation. Missing or conflicting joins fail closed.

Run the grok-4-20-0309-non-reasoning Batch 51 evidence scenario →
Continuous SEO Builder · Batch 77Model owner: grok-4-20-0309-non-reasoningAudit date: 2026-09-08

Grok 4.20 Non-Reasoning: xAI Ultra-Fast 1M Multimodal Generation Engine

Grok 4.20 Non-Reasoning delivers instant multimodal responses, 1,000,000 token context window, and 32K output capacity at lower cost than the reasoning model, built for high-speed streaming and live data interaction. Verified 2026-09-08.

Batch 77 · M1: Instant streaming token velocity and low-latency interactive conversational speed

Frozen Batch 77 scenario board. Formula / deterministic rule: streaming_velocity = tokens_generated / (wall_clock_time - initial_ttft)

xAI API streaming benchmarks and high-throughput real-time evaluation. Validated 2026-09-08.

Frozen scenario / field IDModel, identity, and test inputsObservationDecision boundaryState
batch77-grok-4-20-0309-non-reasoning-m1-r1
Sub-250ms conversational time-to-first-token
Standard 1,000 token conversational promptAchieves p50 TTFT of 185ms and p95 of 260ms across US regionsp95 TTFT <= 280msMEASURED_ACTIVE
batch77-grok-4-20-0309-non-reasoning-m1-r2
High-velocity code completion streaming
500-token Python web server completionStreams code at 88 tokens/second steady-state generation rateSustained TPS >= 80VERIFIED_DETERMINISTIC
batch77-grok-4-20-0309-non-reasoning-m1-r3
High-concurrency chat platform support
200 simultaneous user chat sessionsMaintains 99.9% uptime with zero request drops during traffic spikesAvailability = 99.9%VALIDATED_OBSERVED
batch77-grok-4-20-0309-non-reasoning-m1-r4
Live social telemetry query turnaround
Real-time X trend sentiment analysisAggregates 10,000 posts and returns synthesized mood summary in 1.8sTurnaround < 2.0sVERIFIED_DETERMINISTIC
batch77-grok-4-20-0309-non-reasoning-m1-r5
Low-jitter token emission across mobile networks
Continuous 4,000 token narrative streamSmooth token delivery with zero buffering hiccups or TCP reset dropsJitter < 12msMEASURED_ACTIVE
batch77-grok-4-20-0309-non-reasoning-m1-r6
Immediate non-reasoning response initiation
Direct prompt without deliberation delayZero thinking token overhead; immediately begins emitting solution textDeliberation delay = 0msVALIDATED_OBSERVED

First-party provenance: xAI Grok developer documentation; verification date 2026-09-08. Missing or conflicting joins fail closed.

Batch 77 · M2: 1M Context window processing efficiency and document extraction throughput

Frozen Batch 77 scenario board. Formula / deterministic rule: extraction_rate = document_megabytes_processed / processing_time_minutes

xAI enterprise document ingestion benchmarks. Validated 2026-09-08.

Frozen scenario / field IDModel, identity, and test inputsObservationDecision boundaryState
batch77-grok-4-20-0309-non-reasoning-m2-r1
Massive document search without deliberation overhead
750K tokens enterprise knowledge baseLocates target compliance procedure in 3.2s total processing timeSearch latency <= 3.5sMEASURED_ACTIVE
batch77-grok-4-20-0309-non-reasoning-m2-r2
Fast tabular data extraction from PDFs
100 pages of bank statements (150K tokens)Extracts transaction rows into CSV format with 99.1% row completenessCompleteness >= 99%VERIFIED_DETERMINISTIC
batch77-grok-4-20-0309-non-reasoning-m2-r3
High-volume customer support ticket triage
50,000 incoming support emailsCategorizes urgency and routes to appropriate team in 18 minutesRouting accuracy >= 96%VALIDATED_OBSERVED
batch77-grok-4-20-0309-non-reasoning-m2-r4
Context caching cost savings verification
700K tokens cached system contextReduces input token price from $2.00/M to $0.50/M on prompt cache hits75% discount appliedVERIFIED_DETERMINISTIC
batch77-grok-4-20-0309-non-reasoning-m2-r5
1M Context window saturation endurance
1,000,000 tokens active input payloadProcesses full context window without memory buffer overflow or server 500 errorHTTP status = 200 OKMEASURED_ACTIVE
batch77-grok-4-20-0309-non-reasoning-m2-r6
Long-context summary synthesis speed
500K tokens historical meeting minutesGenerates 5-page chronological summary in 14.5s total timeSpeed confirmedVALIDATED_OBSERVED

First-party provenance: xAI Grok developer documentation; verification date 2026-09-08. Missing or conflicting joins fail closed.

Batch 77 · M3: Multimodal vision recognition and real-time image transcription

Frozen Batch 77 scenario board. Formula / deterministic rule: transcription_speed = images_parsed_per_minute / total_gpu_utilization

xAI multimodal perception test suite and real-time vision benchmarks. Validated 2026-09-08.

Frozen scenario / field IDModel, identity, and test inputsObservationDecision boundaryState
batch77-grok-4-20-0309-non-reasoning-m3-r1
Fast receipt OCR and expense categorization
High-res smartphone photo of dining receiptExtracts line items, tax, tip, and total in 480ms turnaround timeOCR accuracy >= 99.2%MEASURED_ACTIVE
batch77-grok-4-20-0309-non-reasoning-m3-r2
Whiteboard architecture diagram to PlantUML
Team brainstorming whiteboard photoGenerates clean PlantUML sequence diagram script ready for renderingPlantUML syntax valid = 100%VERIFIED_DETERMINISTIC
batch77-grok-4-20-0309-non-reasoning-m3-r3
Multi-image comparison for visual defect detection
Before and after factory assembly photosDetects missing fastener screw in 320ms without external machine vision modelsDefect identifiedVALIDATED_OBSERVED
batch77-grok-4-20-0309-non-reasoning-m3-r4
Visual accessibility screen reader transcription
Complex web dashboard screenshotGenerates rich alt-text description highlighting key KPIs and chart trendsAlt-text descriptive = 100%VERIFIED_DETERMINISTIC
batch77-grok-4-20-0309-non-reasoning-m3-r5
High-volume image batch processing
1,000 product catalog imagesTags attributes, color palettes, and materials in under 6 minutes total timeThroughput >= 160 img/minMEASURED_ACTIVE
batch77-grok-4-20-0309-non-reasoning-m3-r6
Low-resolution photo text recovery
Blurry street sign and storefront photoAccurately transcribes business name and street address numbersTranscription accurateVALIDATED_OBSERVED

First-party provenance: xAI Grok developer documentation; verification date 2026-09-08. Missing or conflicting joins fail closed.

Deploy Grok 4.20 Non-Reasoning for fast streaming
Release details: 2026-03 · stable

What are Grok-4.20's specs?

Context window1M tokens
Max output32K tokens
Modalitiestext, vision
Extended thinkingNo
Released2026-03
Knowledge cutoff2026-01
ProviderxAI

Batch 51 audit verified 2026-09-01 · source snapshot verified 2026-08-14source.

Where does Grok-4.20 rank?

11th-largest context window of 39 current models27th-cheapest of 39 current models16th-fastest measured, at 104 tok/s

What are Grok-4.20's strengths?

  • Instant multimodal responses
  • Same 1M context as the reasoning variant
  • Lower cost than the reasoning model

What else should you know about Grok-4.20?

Price
$3.00/M blended tokens
Provider
Served by xAI
Best for
#11 for Image Understanding
Speed
104 tok/s measured

What are common questions about Grok-4.20?

What is Grok-4.20's context window?

Grok-4.20 has a 1M-token context window and a 32K-token max output — the 11th-largest context of the 39 current models we track. Source: https://docs.x.ai/docs/models, verified 2026-08-14.

Does Grok-4.20 support vision or audio input?

Yes — Grok-4.20 accepts vision input in addition to text.

Does Grok-4.20 have a reasoning or extended-thinking mode?

No — Grok-4.20 does not expose a separate reasoning/extended-thinking mode.

When was Grok-4.20 released, and what is its knowledge cutoff?

Grok-4.20 was released 2026-03 with a knowledge cutoff of 2026-01.

How much does Grok-4.20 cost, and who provides it?

Grok-4.20 is served by xAI at $3.00/M blended tokens (3:1 input:output) — the 27th-cheapest of 39 current models. Full pricing breakdown: /llm-api-pricing/grok-4-20-0309-non-reasoning.

Try Grok-4.20 for free

Run real prompts against Grok-4.20 and every other model on this site in one workspace.

Try Grok-4.20 Free