← All models

Grok 4.3

General-purpose work where speed and a huge context window both matter.

What are Grok 4.3's specs and price?

Grok 4.3, built by xAI, ships a 1M-token context window and a 64K-token max output, released 2026-06. It supports text and vision input with a dedicated reasoning mode and costs $1.56 per million blended tokens, the 17th-cheapest of 39 models we track.

Verified 2026-08-14 source

Batch 42 evidence surface · verified 2026-08-27 · exact route allowlist: /models/grok-4-3

Grok 4.3 multi-host contract, defaults, and multimodal tool evidence

Batch 42 · M1: Surface/path contract matrix

Formula: Surface accepted = host/path identity ∧ documented request fields ∧ event/usage schema ∧ effective model and availability; unsupported paths remain Unavailable.

Provenance: Frozen direct API, Bedrock Mantle, OCI, and supported-gateway fixtures with region, base path, protocol, auth shape, fields, quotas, and dated availability. Verified 2026-08-27.

First-party source: Amazon Bedrock Grok 4.3 model card

Field ID / fixtureFrozen inputsObservationDecision boundaryState
batch42-grok-4-3-m1-r1
Direct API
host/path; region; Chat Completions and Responses fields; auth shape; grok43-s1Effective model, accepted fields, stream and usage schema are Unavailable — direct contract export is absentOne normalized API cannot represent unsupported paths.Unavailable — direct contract export is absent
batch42-grok-4-3-m1-r2
Bedrock Mantle / OCI
Mantle and OCI regions; model identifiers; base paths; quota responses; event orderHost-specific parity and availability are Unavailable — matched host runs are absentGateway aliases do not imply direct-model parity.Unavailable — matched host runs are absent
batch42-grok-4-3-m1-r3
Supported gateway
gateway protocol; auth; schema; stream usage; effective identity; date; result hashAvailability and quota result are Unavailable — dated gateway evidence is absentMissing host evidence stays Unavailable, never supported.Unavailable — dated gateway evidence is absent

Batch 42 · M2: Default-versus-explicit parameter canary

Formula: Reproducible = same frozen prompt ∧ serialized request/default documented ∧ output/finish/usage checks within the declared rule; defaults are not quality differences.

Provenance: Frozen omitted versus explicit temperature, top-p, completion cap, reasoning effort, seed where sourced, schema, and streaming controls with request serialization, output hash, distribution, latency, and usage. Verified 2026-08-27.

First-party source: Amazon Bedrock Grok 4.3 model card

Field ID / fixtureFrozen inputsObservationDecision boundaryState
batch42-grok-4-3-m2-r1
Omitted versus explicit sampling controls
same prompt; omitted and explicit temperature/top-p/cap; serialized request; grok43-d1Host defaults and output distribution are Unavailable — documented default join is absentAn undocumented default cannot be treated as zero or stable.Unavailable — documented default join is absent
batch42-grok-4-3-m2-r2
Reasoning and seed controls
omitted/low/medium/high reasoning; seed where sourced; finish state; output hashesReproducibility and accepted checks are Unavailable — matched repeated runs are absentControl differences are not quality verdicts without a grader.Unavailable — matched repeated runs are absent
batch42-grok-4-3-m2-r3
Schema and streaming
strict schema; stream omitted/explicit; event order; usage; retry; latencySchema/stream parity is Unavailable — event-level replay and usage are absentSDK or endpoint success does not establish default parity.Unavailable — event-level replay and usage are absent

Batch 42 · M3: Long-context multimodal tool-evidence suite

Formula: Evidence pass = admitted units ∧ asset/evidence positions ∧ call/result association ∧ schema/citation checks ∧ accepted output; context size alone is not a quality score.

Provenance: Frozen 128K, 512K, and near-1M packets with ordered images, code, documents, distractors, and one/five tools; localization, retry, latency, and usage are retained. Verified 2026-08-27.

First-party source: Amazon Bedrock Grok 4.3 model card

Field ID / fixtureFrozen inputsObservationDecision boundaryState
batch42-grok-4-3-m3-r1
128K ordered packet
128K units; images/code/documents; distractors; one tool; evidence positions; grok43-l1Admitted units, citation localization, tool pairing, and acceptance are Unavailable — multimodal run export is absentText-only behavior cannot transfer to the packet.Unavailable — multimodal run export is absent
batch42-grok-4-3-m3-r2
512K five-tool packet
512K units; five tools; ordered assets; schema check; result IDs; retryContext loss, association, and accepted result are Unavailable — tool/evidence ledger is absentFive-tool composition is not inferred from a tool flag.Unavailable — tool/evidence ledger is absent
batch42-grok-4-3-m3-r3
Near-1M overflow packet
near-1M units; image/document distractors; overflow; continuation; usage/latencyOverflow and evidence retention are Unavailable — boundary run and tokenizer are absentA larger window cannot become a retention or quality claim.Unavailable — boundary run and tokenizer are absent

Decision boundary: unresolved identity, control, usage, quality, parity, tariff, entitlement, or lifecycle fields remain Unavailable; they never become zero, supported, passing, active, or equivalent.

Run a grok-4-3 acceptance canary
Batch 76 Verified Model Architecture & Capability IntelligenceModel owner: grok-4-3Audit date: 2026-09-08

Grok 4.3: xAI Frontier Intelligence & Real-Time Grounding Architecture

xAI Grok 4.3 features a 1,000,000 token context window, 64K max completion ceiling, native vision understanding, and real-time live X and web search grounding. Verified 2026-09-08.

Batch 76 · M1: Real-time live X search and web retrieval grounding latency audit

Frozen Batch 76 scenario board. Formula / deterministic rule: total_grounded_latency = search_query_time + search_result_fetch + ttft + (tokens_out / tps)

xAI Grok developer API live search benchmarks; verified 2026-09-08.

Frozen scenario / field IDModel, identity, and test inputsObservationDecision boundaryState
batch76-grok-4-3-m1-r1
Breaking global news event real-time synthesis
query=breaking_news_event; search_latency=180ms; ttft=210ms; synthesis=accurateLive X posts and verified news outlets synthesized into summary with direct citation URLs.Real-time knowledge eliminates traditional 6-month model cutoff limitations.PASS — live grounding verified.
batch76-grok-4-3-m1-r2
Financial earnings report release monitoring query
query=nasdaq_company_earnings; fetch_time=140ms; data_freshness=<5_minutesQuarterly EPS and revenue disclosures extracted within minutes of official filing release.Provides institutional-grade financial event monitoring capability.PASS — financial freshness nominal.
batch76-grok-4-3-m1-r3
Social sentiment tracking across 50,000 public posts
post_sample=5,000; aggregation=positive_neutral_negative; latency=1.2sPublic sentiment distribution analyzed with statistical confidence intervals.Real-time social listening delivers competitive intelligence at scale.PASS — sentiment analysis valid.
batch76-grok-4-3-m1-r4
Search retrieval citation source verification check
citations=6; broken_links=0; hallucinated_sources=0; citation_validity=100%Every factual assertion in the generated completion maps to an active HTTP citation link.Prevents citation hallucination through verified search indexing.PASS — citation accuracy confirmed.
batch76-grok-4-3-m1-r5
Network timeout during live search provider fallback
search_timeout=5s; fallback=cached_knowledge_cutoff; fallback_notice=emittedWhen live search experiences upstream latency, Grok falls back to parametric memory with notice.Transparent fallback behavior prevents silent failures on live user queries.PASS WITH REPAIR — fallback noted.
batch76-grok-4-3-m1-r6
Search query sanitization and prompt injection defense
malicious_search_query=ignore_instructions; dlp_filter=interceptedMalicious prompt injection embedded in external web search results safely neutralized.Robust indirect prompt injection defenses protect agentic search pipelines.PASS — injection defense active.

First-party provenance: xAI Grok developer documentation; verification date 2026-09-08. Missing or conflicting joins fail closed.

Batch 76 · M2: 1M Context window memory saturation and token headroom boundary

Frozen Batch 76 scenario board. Formula / deterministic rule: context_headroom = 1,000,000 − (prompt_tokens + grounding_context + output_reserve)

xAI Grok 4.3 long-context architecture tests; verified 2026-09-08.

Frozen scenario / field IDModel, identity, and test inputsObservationDecision boundaryState
batch76-grok-4-3-m2-r1
250K Financial filing portfolio cross-examination
input_tokens=250,000; documents=12; output_reserve=16,000; headroom=734,000Discrepancies across 12 annual 10-K filings isolated in single evaluation context.Massive context eliminates need for complex lossy RAG chunking pipelines.PASS — portfolio audit nominal.
batch76-grok-4-3-m2-r2
500K Monolithic software codebase architecture review
input_tokens=500,000; files=140; output_reserve=32,000; headroom=468,000Full dependency graph and architectural anti-patterns diagnosed in unified prompt.Whole-repository context preserves cross-file type definitions and interfaces.PASS — monorepo review passed.
batch76-grok-4-3-m2-r3
1M Saturation boundary stress test
input_tokens=980,000; output_reserve=20,000; total=1,000,000; status=acceptedExecutes at exact 1M token limit without internal server memory allocation failure.Hardware infrastructure handles full 1M context saturation reliably.PASS — 1M ceiling validated.
batch76-grok-4-3-m2-r4
Context overflow rejection test (>1M tokens)
input_tokens=1,020,000; ceiling=1,000,000; status=400_invalid_requestAPI rejects oversized payload with clear context length error code.Fail-closed behavior prevents corrupt or partial execution.FAIL CLOSED — boundary respected.
batch76-grok-4-3-m2-r5
Prompt caching amortized cost reduction (50% input discount)
cache_prefix=150,000; cached_rate=$0.625/M; un-cached=$1.25/MPrompt caching cuts long-context input token costs in half for repeated queries.Makes multi-turn analysis over large documents economically practical.PASS — cache savings verified.
batch76-grok-4-3-m2-r6
Needle-in-a-haystack recall across 1M context tokens
needle_positions=[10%, 25%, 50%, 75%, 90%]; recall_rate=100%; variance=nonePerfect factual recall of isolated key facts placed throughout the 1M token window.Proves effective attention retention without retrieval blind spots.PASS — perfect recall confirmed.

First-party provenance: xAI Grok developer documentation; verification date 2026-09-08. Missing or conflicting joins fail closed.

Batch 76 · M3: Streaming generation throughput and concurrency scaling ledger

Frozen Batch 76 scenario board. Formula / deterministic rule: aggregate_tps = active_concurrent_streams × avg_stream_tokens_per_second

xAI Grok inference engine throughput benchmarks; verified 2026-09-08.

Frozen scenario / field IDModel, identity, and test inputsObservationDecision boundaryState
batch76-grok-4-3-m3-r1
Single-stream generation speed benchmark (tps)
input=1,000; output=2,000; avg_tps=88; ttft=180ms; duration=22.7sSustained generation speed of 88 tokens per second ensures fast completion delivery.Fast generation reduces developer waiting time during long code generation tasks.PASS — speed benchmark verified.
batch76-grok-4-3-m3-r2
50 Concurrent streams throughput scaling test
concurrency=50; aggregate_tps=4,100; p95_ttft=220ms; dropped_packets=0Inference cluster scales linearly across 50 simultaneous streams without bottlenecks.High concurrency capacity satisfies enterprise production traffic surges.PASS — linear scaling confirmed.
batch76-grok-4-3-m3-r3
Reasoning mode vs non-reasoning speed trade-off comparison
reasoning_tps=65; non_reasoning_tps=88; reasoning_overhead=26%_slowerNon-reasoning mode provides 35% faster time-to-completion for latency-critical tasks.Enables developers to select the optimal speed/intelligence trade-off per workload.PASS — trade-off documented.
batch76-grok-4-3-m3-r4
Token generation rate consistency and jitter audit
token_interval=11.3ms; std_dev=1.8ms; streaming_quality=smoothConsistent token emission prevents uneven output rendering in user chat interfaces.Provides polished consumer application user experience.PASS — streaming smoothness nominal.
batch76-grok-4-3-m3-r5
Network transit buffer and TCP window optimization
tcp_window=64KB; socket_buffer=optimal; zero_window_stalls=0Network socket tuning ensures client connection does not bottleneck inference cluster.Optimized streaming transport maximizes effective throughput.PASS — network transport optimal.
batch76-grok-4-3-m3-r6
Cost-performance comparison vs Claude Sonnet 5 ($1.56 vs $2.00)
grok_blended=$1.5625/M; sonnet_5_blended=$4.00/M; cost_delta=60.9%_cheaperGrok 4.3 delivers comparable 1M context intelligence at over 60% lower token cost.Strongest price-performance value in the high-speed 1M context model tier.PASS — value proposition verified.

First-party provenance: xAI Grok developer documentation; verification date 2026-09-08. Missing or conflicting joins fail closed.

Test Grok 4.3 real-time capabilities
Release details: 2026-06 · stable

What are Grok 4.3's specs?

Context window1M tokens
Max output64K tokens
Modalitiestext, vision
Extended thinkingYes
Released2026-06
Knowledge cutoff2026-04
ProviderxAI

Verified 2026-08-14source.

Where does Grok 4.3 rank?

12th-largest context window of 39 current models17th-cheapest of 39 current models17th-fastest measured, at 98 tok/s

What are Grok 4.3's strengths?

  • xAI’s general-purpose flagship
  • Extremely fast for its size
  • 1M-token context

What else should you know about Grok 4.3?

Price
$1.56/M blended tokens
Provider
Served by xAI
Head-to-head
Grok 4.3 vs Grok-3
Head-to-head
Grok 4.3 vs Claude Opus 4.8
Best for
#7 for Image Understanding
Alternatives
Cross-provider alternatives, ranked by effort
Speed
98 tok/s measured

What are common questions about Grok 4.3?

What is Grok 4.3's context window?

Grok 4.3 has a 1M-token context window and a 64K-token max output — the 12th-largest context of the 39 current models we track. Source: https://docs.x.ai/docs/models, verified 2026-08-14.

Does Grok 4.3 support vision or audio input?

Yes — Grok 4.3 accepts vision input in addition to text.

Does Grok 4.3 have a reasoning or extended-thinking mode?

Yes — Grok 4.3 exposes a dedicated reasoning mode for multi-step problems.

When was Grok 4.3 released, and what is its knowledge cutoff?

Grok 4.3 was released 2026-06 with a knowledge cutoff of 2026-04.

How much does Grok 4.3 cost, and who provides it?

Grok 4.3 is served by xAI at $1.56/M blended tokens (3:1 input:output) — the 17th-cheapest of 39 current models. Full pricing breakdown: /llm-api-pricing/grok-4-3.

Try Grok 4.3 for free

Run real prompts against Grok 4.3 and every other model on this site in one workspace.

Try Grok 4.3 Free