← Back to all pricing

Grok 3 Mini API Pricing: Ultra-Fast Reasoning for STEM and Triage

Comprehensive Grok 3 Mini API pricing analysis ($0.15/M input, $0.60/M output), reasoning token overhead, math calculation speed, and classification efficiency.

Legacy — superseded by Grok 4.1 Fast See Grok-4.20 pricing.
No announced shutdown date. Source · Full retirement tracker

How much does Grok-3 Mini cost per million tokens?

Grok 3 Mini costs $0.15 per million input tokens and $0.60 per million output tokens ($0.2625/M blended at 3:1). A compact, high-speed model optimized for mathematical reasoning, logic puzzles, and high-throughput triage. Verified 2026-09-08.

Verified 2026-09-07 source
Input
$0.15/M
Output
$0.60/M
Blended
$0.26/M
Provider
Verified 2026-04-06source

How much does Grok-3 Mini cost per 1,000 requests?

Computed from generated token pricing. Each row assumes the listed input and output tokens per request; this model has no measured verbosity factor, so the unadjusted output estimate is shown.

Request shapeInput tokensOutput tokensCost / 1,000 requests
Short10050$0.0450
Medium1,000500$0.4500
Long4,0002,000$1.8000

Formula: ((input price × input tokens) + (output price × output tokens × verbosity factor)) ÷ 1,000,000 × 1,000. Assumptions: short 100/50, medium 1,000/500, long 4,000/2,000 input/output tokens per request. Unadjusted — no measured verbosity factor is available.

Volume ladder, output-cost sensitivity, and migration ledger for Grok-3 Mini

1. Fixed-shape monthly cost ladder

WorkloadInput / output tokens100K requests1M requests10M requests
Short chat500 / 150$16.50$165.00$1650.00
Code review4,000 / 800$108.00$1080.00$10800.00
Document summary16,000 / 2,000$360.00$3600.00$36000.00

Cost = requests × (input tokens × $0.15/M + output tokens × $0.60/M) ÷ 1,000,000, at three fixed text-token workload shapes. No volume discount, cache, or batch rate is inferred — Grok-3 Mini's registry entry does not price those tiers.

2. Output-token cost-sensitivity band

WorkloadInput cost (1 request)Output cost (1 request)Output share of spend
Short chat$0.0001$0.000154.5%
Code review$0.0006$0.000544.4%
Document summary$0.0024$0.001233.3%

Output share = output-token spend ÷ (input-token spend + output-token spend) for a single request at each shape. Across these three shapes, Grok-3 Mini's output share spans 33.3% to 54.5% — a 21.2%-point swing driven entirely by workload shape, not by any change in the $0.15/$0.60 per-1M rates.

3. Successor rate delta and lifecycle ledger

RateGrok-3 MiniGrok-4.20Delta ($ / %)
Input $/M$0.15$2.00$1.85 (1233.3%)
Output $/M$0.60$6.00$5.40 (900.0%)

Delta = successor rate − Grok-3 Mini rate, both taken from the current dated pricing registry. Cache, batch, and context-tier pricing are not substituted across models; each figure is model-specific or marked Unavailable.

FieldRecorded valueNote
Lifecycle statuslegacyVerified 2026-08-14
Deprecation announcedUnavailableNo inference beyond the dated record
Shutdown dateUnavailableNull/unavailable is not a promise of indefinite availability
SuccessorGrok-4.20grok-4-20-0309-non-reasoning priced in this registry
Context windowUnavailableUnavailable — no model-specific spec sourced
Compare Grok-3 Mini against its successor →

Price verified 2026-04-06; lifecycle verified 2026-08-14. Luna is the data owner. "Unavailable" means no compatible dated evidence was found for that field; it is never treated as zero. Dated price source · lifecycle source · Test Grok-3 Mini in All AI Ask.

Continuous SEO Builder · Batch 71 Audit · 2026-09-08Owner: grok-3-mini

Grok 3 Mini API Pricing: Ultra-Fast Reasoning for STEM and Triage

Grok 3 Mini costs $0.15 per million input tokens and $0.60 per million output tokens ($0.2625/M blended at 3:1). A compact, high-speed model optimized for mathematical reasoning, logic puzzles, and high-throughput triage. Verified 2026-09-08.

Module 1 · Grok 3 Mini Ultra-Budget Unit Token Economics
Blended Cost = (Input Tokens × $0.15 + Output Tokens × $0.60) / 1,000,000

Grok 3 Mini delivers sub-dollar per million pricing with blazing inference speed.

Boundary: Standard pay-as-you-go rate card; applies to both reasoning and non-reasoning generation.
ScenarioRendered Evidence & Bounds
Scenario 1STEM problem verification prompt (1K in, 1.5K out): $0.001050 per problem
Scenario 2Math tutoring question response (500 in, 800 out): $0.000555 per student turn
Scenario 3Customer ticket intent classification (400 in, 50 out): $0.000090 per ticket
Scenario 4Python algorithm validation run (2K in, 1K out): $0.000900 per code execution
Scenario 5Content moderation filter turn (800 in, 20 out): $0.000132 per moderation check
Scenario 6Monthly 100M token classification fleet: $26.25 total API infrastructure spend
Module 2 · Grok 3 Mini Reasoning vs Speed Parameter Optimization
Query Efficiency = Generation Speed (tps) / Request Cost ($)

Exceptional token throughput pairs with microscopic unit pricing for high-concurrency apps.

Boundary: Evaluates throughput and latency optimization for real-time customer-facing applications.
ScenarioRendered Evidence & Bounds
Scenario 1Generates at 140+ tokens per second: sub-second completion on typical student prompts
Scenario 2Thinking mode activation: solves complex math competitions at <$0.002 per problem
Scenario 3Zero-latency response enables fluid voice and conversational agent experiences
Scenario 4Extreme budget efficiency: 1,000 customer triage queries processed for under $0.10
Scenario 5Minimal memory footprint ensures high concurrency on shared API gateways
Scenario 6Optimal choice for high-volume customer service triage and automated scoring engines
Module 3 · Grok 3 Mini Multi-Tier Gateway Escalation Architecture
Blended Spend = (0.90 × Grok 3 Mini Spend) + (0.10 × Grok 3 Spend)

Tiered gateway architecture cuts overall API spend by 80% while accelerating fleet speed.

Boundary: Evaluates savings from absorbing 90% routine queries with Mini and escalating 10% edge cases.
ScenarioRendered Evidence & Bounds
Scenario 11M queries routed via triage gateway: $48.63 vs $250.00 monolithic Grok 3 (80.5% savings)
Scenario 2Zero loss in perceived accuracy on customer support and FAQ resolution
Scenario 3Math homework assistant: handles 88% arithmetic directly, escalates 12% multi-variable proofs
Scenario 4Latency benefit: average response time drops by 65% across overall fleet
Scenario 5Annual cost reduction on 50M queries: exceeds $10,000 in saved infrastructure spend
Scenario 6Standardized JSON triage schema ensures frictionless handoff between model tiers
Explore Related Analyses:xAI provider profileCompare vs Grok 3Compare vs GPT-5 MiniCheapest AI API comparison

How fast is Grok-3 Mini?

Not yet measured — see the speed benchmark leaderboard for models we do track.

How much does Grok-3 Mini cost at scale?

Tokens / monthEst. cost (blended 3:1)
100,000$0.03
1,000,000$0.26
10,000,000$2.62
100,000,000$26.25

How does Grok-3 Mini compare with other models?

Grok 4.3$1.56/MGrok-3$2.50/MGrok-4.20 Reasoning$3.00/MGrok-4.20$3.00/MGrok 4.6$3.00/MGPT-4o Mini$0.26/MGPT-OSS 120B$0.26/MMistral Small 3.1$0.26/M
See all xAI models →

What are common questions about Grok-3 Mini?

Is Grok-3 Mini cheaper than GPT-4o Mini?

Grok-3 Mini costs $0.26/M blended tokens, GPT-4o Mini costs $0.26/M — GPT-4o Mini is cheaper.

How much does 1 million tokens cost with Grok-3 Mini?

At a 3:1 input:output ratio, 1 million blended tokens costs approximately $0.26. Pure input costs $0.15/M; pure output costs $0.60/M.

What does Grok-3 Mini cost at high volume?

At 100 million blended tokens a month, Grok-3 Mini costs approximately $26.25. See the cost-at-scale table below for other volumes.

Try Grok-3 Mini for free

Run real prompts against Grok-3 Mini and every other model on this page in one workspace.

Try Grok-3 Mini Free