Mistral Small API Pricing: European Enterprise Cost Efficiency
Explore Mistral Small API pricing ($0.15/M input, $0.60/M output), European GDPR data sovereignty, fast instruction following, and enterprise API rates.
Full specs, context window and API limits →How much does Mistral Small 3.1 cost per million tokens?
Mistral Small costs $0.15 per million input tokens and $0.60 per million output tokens ($0.2625/M blended at 3:1). Provides lightweight European sovereign AI with 128K context and rapid instruction following. Verified 2026-09-08.
How much does Mistral Small 3.1 cost per 1,000 requests?
Computed from generated token pricing. Each row assumes the listed input and output tokens per request; output is adjusted by this model's measured 0.85× verbosity factor.
| Request shape | Input tokens | Output tokens | Cost / 1,000 requests |
|---|---|---|---|
| Short | 100 | 50 | $0.0405 |
| Medium | 1,000 | 500 | $0.4050 |
| Long | 4,000 | 2,000 | $1.6200 |
Formula: ((input price × input tokens) + (output price × output tokens × verbosity factor)) ÷ 1,000,000 × 1,000. Assumptions: short 100/50, medium 1,000/500, long 4,000/2,000 input/output tokens per request. Verbosity run: 2026-06-16T20:31:30.728Z.
Three model-specific pricing decisions
Small owns everyday extraction, chat, and vision-shaped decision economics. Image units are not converted from text tokens.
1. Extraction, chat, and vision-shaped bills
| Workload | Fixed shape | 100K bill | Image-unit boundary |
|---|---|---|---|
| Extraction | 2,000 / 400 | $54.00 | Text tokens |
| Chat | 1,000 / 500 | $45.00 | Text tokens |
| Vision-shaped | 2,000 text + image | $54.00 | Image unit unavailable; text row is not an image bill |
2. Small versus Ministral accepted-result crossover
| Fixed input / output | Mistral Small 3.1 | Ministral 8B | Narrow decision boundary |
|---|---|---|---|
| Short task · 1,000 / 300 | $33.00 | $19.50 | 10% accepted-result uplift required |
| 256K-shaped task · 256,000 / 2,000 | $3960.00 | $3870.00 | 15% accepted-result uplift required |
Formula: requests × (input tokens × input $/M + output tokens × output $/M) ÷ 1,000,000. The uplift is a planning threshold, not a measured quality claim.
3. Dated family price / TTFT / throughput ordering
| Model | Verified rates | Speed evidence | Ordering boundary |
|---|---|---|---|
| Mistral Small 3.1 | 2026-08-14 · $0.15 / $0.60 per M | 121 tokens/sec; TTFT 260 ms; 5 measured samples | Price/speed input only; no family verdict |
| Mistral Large 3 | 2026-08-14 · $0.50 / $1.50 per M | 61 tokens/sec; TTFT 400 ms; 5 measured samples | Price/speed input only; no family verdict |
| Mistral Medium 3 | 2026-08-14 · $1.50 / $7.50 per M | 92 tokens/sec; TTFT 320 ms; 5 measured samples | Price/speed input only; no family verdict |
Verified 2026-08-14. Luna is the data owner. “Unavailable” means no compatible dated evidence was found; it is never treated as zero. First-party price source · Run this Batch 5 scenario.
All three Batch 5 decisions are server-rendered for Mistral Small 3.1; fixed inputs, formulas, dated sources, speed sample state, and unavailable mechanics are visible.
Batch 68 · exact model pricing decision contributions · verified 2026-09-08
Exact model boundary: Mistral mistral-small (slug mistral-small). First-party provider pricing and API documentation remain fact owners.
European sovereign token pricing and monthly spend matrix
Frozen Batch 68 scenario board. Formula / deterministic rule: monthly_spend = calls * ((in_tokens * 0.15 + out_tokens * 0.60) / 1M) Boundary: Owns Mistral Small base token tariffs and European commercial billing.
| Frozen scenario / field ID | Model, identity, provider, and evidence fields | Result | State |
|---|---|---|---|
batch68-mistral-small-m1-r150K multilingual customer support tickets | model=mistral-small; slug=mistral-small; provider=Mistral; scenario=50K multilingual customer support tickets; workload; prompt tokens; completion tokens; standard spend; batch spend; cost winner; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-08; measurement versus assumption=explicit | Unavailable — 50K multilingual customer support tickets is a frozen fixture pending exact identity, configuration, and denominator joins. | UNTESTED — assumption cannot establish a verdict |
batch68-mistral-small-m1-r2200K structured JSON entity extractions | model=mistral-small; slug=mistral-small; provider=Mistral; scenario=200K structured JSON entity extractions; workload; prompt tokens; completion tokens; standard spend; batch spend; cost winner; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-08; measurement versus assumption=explicit | Unavailable — 200K structured JSON entity extractions is a frozen fixture pending exact identity, configuration, and denominator joins. | UNTESTED — assumption cannot establish a verdict |
batch68-mistral-small-m1-r31M high-volume GDPR classification events | model=mistral-small; slug=mistral-small; provider=Mistral; scenario=1M high-volume GDPR classification events; workload; prompt tokens; completion tokens; standard spend; batch spend; cost winner; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-08; measurement versus assumption=explicit | Unavailable — 1M high-volume GDPR classification events is a frozen fixture pending exact identity, configuration, and denominator joins. | UNTESTED — assumption cannot establish a verdict |
batch68-mistral-small-m1-r4batch API queue (50% discount) | model=mistral-small; slug=mistral-small; provider=Mistral; scenario=batch API queue (50% discount); workload; prompt tokens; completion tokens; standard spend; batch spend; cost winner; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-08; measurement versus assumption=explicit | Unavailable — batch API queue (50% discount) is a frozen fixture pending exact identity, configuration, and denominator joins. | UNTESTED — assumption cannot establish a verdict |
batch68-mistral-small-m1-r5EU-only dedicated instance deployment | model=mistral-small; slug=mistral-small; provider=Mistral; scenario=EU-only dedicated instance deployment; workload; prompt tokens; completion tokens; standard spend; batch spend; cost winner; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-08; measurement versus assumption=explicit | Unavailable — EU-only dedicated instance deployment is a frozen fixture pending exact identity, configuration, and denominator joins. | UNTESTED — assumption cannot establish a verdict |
batch68-mistral-small-m1-r6unresolved billing currency | model=mistral-small; slug=mistral-small; provider=Mistral; scenario=unresolved billing currency; workload; prompt tokens; completion tokens; standard spend; batch spend; cost winner; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-08; measurement versus assumption=explicit | Unavailable — unresolved billing currency has no matched, dated bilateral observation. | FAIL CLOSED — manual, probe, or source evidence required |
First-party provenance: Mistral AI model pricing; verification date 2026-09-08. Missing or conflicting joins fail closed.
Multilingual translation and European regulatory compliance gate
Frozen Batch 68 scenario board. Formula / deterministic rule: compliance_roi = regulatory_assurance_value - ((in * 0.15 + out * 0.60) / 1M) Boundary: Owns GDPR data residency and multilingual European language accuracy.
| Frozen scenario / field ID | Model, identity, provider, and evidence fields | Result | State |
|---|---|---|---|
batch68-mistral-small-m2-r1French legal contract summary | model=mistral-small; slug=mistral-small; provider=Mistral; scenario=French legal contract summary; language pair; document size; translation fidelity; token expenditure; sovereignty verdict; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-08; measurement versus assumption=explicit | Unavailable — French legal contract summary is a frozen fixture pending exact identity, configuration, and denominator joins. | UNTESTED — assumption cannot establish a verdict |
batch68-mistral-small-m2-r2German technical documentation translation | model=mistral-small; slug=mistral-small; provider=Mistral; scenario=German technical documentation translation; language pair; document size; translation fidelity; token expenditure; sovereignty verdict; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-08; measurement versus assumption=explicit | Unavailable — German technical documentation translation is a frozen fixture pending exact identity, configuration, and denominator joins. | UNTESTED — assumption cannot establish a verdict |
batch68-mistral-small-m2-r3Spanish customer privacy inquiry | model=mistral-small; slug=mistral-small; provider=Mistral; scenario=Spanish customer privacy inquiry; language pair; document size; translation fidelity; token expenditure; sovereignty verdict; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-08; measurement versus assumption=explicit | Unavailable — Spanish customer privacy inquiry is a frozen fixture pending exact identity, configuration, and denominator joins. | UNTESTED — assumption cannot establish a verdict |
batch68-mistral-small-m2-r4cross-border multilingual support ticket | model=mistral-small; slug=mistral-small; provider=Mistral; scenario=cross-border multilingual support ticket; language pair; document size; translation fidelity; token expenditure; sovereignty verdict; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-08; measurement versus assumption=explicit | Unavailable — cross-border multilingual support ticket is a frozen fixture pending exact identity, configuration, and denominator joins. | UNTESTED — assumption cannot establish a verdict |
batch68-mistral-small-m2-r5data transfer outside EEA guardrail | model=mistral-small; slug=mistral-small; provider=Mistral; scenario=data transfer outside EEA guardrail; language pair; document size; translation fidelity; token expenditure; sovereignty verdict; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-08; measurement versus assumption=explicit | Unavailable — data transfer outside EEA guardrail is a frozen fixture pending exact identity, configuration, and denominator joins. | UNTESTED — assumption cannot establish a verdict |
batch68-mistral-small-m2-r6unsupported European dialect | model=mistral-small; slug=mistral-small; provider=Mistral; scenario=unsupported European dialect; language pair; document size; translation fidelity; token expenditure; sovereignty verdict; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-08; measurement versus assumption=explicit | Unavailable — unsupported European dialect has no matched, dated bilateral observation. | FAIL CLOSED — manual, probe, or source evidence required |
First-party provenance: Mistral AI documentation; verification date 2026-09-08. Missing or conflicting joins fail closed.
Two-tier enterprise routing: Mistral Small vs Mistral Large
Frozen Batch 68 scenario board. Formula / deterministic rule: blended_cost = small_volume * small_cost + large_volume * large_cost Boundary: Owns two-tier architectural routing between cost-efficient Small and frontier Large.
| Frozen scenario / field ID | Model, identity, provider, and evidence fields | Result | State |
|---|---|---|---|
batch68-mistral-small-m3-r1100% Mistral Small baseline ($0.15/$0.60) | model=mistral-small; slug=mistral-small; provider=Mistral; scenario=100% Mistral Small baseline ($0.15/$0.60); escalation mix; total monthly queries; blended spend; cost savings vs pure Large; quality retention; routing verdict; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-08; measurement versus assumption=explicit | Unavailable — 100% Mistral Small baseline ($0.15/$0.60) is a frozen fixture pending exact identity, configuration, and denominator joins. | UNTESTED — assumption cannot establish a verdict |
batch68-mistral-small-m3-r290% Small triage / 10% Large escalation | model=mistral-small; slug=mistral-small; provider=Mistral; scenario=90% Small triage / 10% Large escalation; escalation mix; total monthly queries; blended spend; cost savings vs pure Large; quality retention; routing verdict; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-08; measurement versus assumption=explicit | Unavailable — 90% Small triage / 10% Large escalation is a frozen fixture pending exact identity, configuration, and denominator joins. | UNTESTED — assumption cannot establish a verdict |
batch68-mistral-small-m3-r380% Small triage / 20% Large escalation | model=mistral-small; slug=mistral-small; provider=Mistral; scenario=80% Small triage / 20% Large escalation; escalation mix; total monthly queries; blended spend; cost savings vs pure Large; quality retention; routing verdict; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-08; measurement versus assumption=explicit | Unavailable — 80% Small triage / 20% Large escalation is a frozen fixture pending exact identity, configuration, and denominator joins. | UNTESTED — assumption cannot establish a verdict |
batch68-mistral-small-m3-r450% Small triage / 50% Large escalation | model=mistral-small; slug=mistral-small; provider=Mistral; scenario=50% Small triage / 50% Large escalation; escalation mix; total monthly queries; blended spend; cost savings vs pure Large; quality retention; routing verdict; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-08; measurement versus assumption=explicit | Unavailable — 50% Small triage / 50% Large escalation is a frozen fixture pending exact identity, configuration, and denominator joins. | UNTESTED — assumption cannot establish a verdict |
batch68-mistral-small-m3-r5100% direct Mistral Large execution | model=mistral-small; slug=mistral-small; provider=Mistral; scenario=100% direct Mistral Large execution; escalation mix; total monthly queries; blended spend; cost savings vs pure Large; quality retention; routing verdict; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-08; measurement versus assumption=explicit | Unavailable — 100% direct Mistral Large execution is a frozen fixture pending exact identity, configuration, and denominator joins. | UNTESTED — assumption cannot establish a verdict |
batch68-mistral-small-m3-r6unresolved routing confidence threshold | model=mistral-small; slug=mistral-small; provider=Mistral; scenario=unresolved routing confidence threshold; escalation mix; total monthly queries; blended spend; cost savings vs pure Large; quality retention; routing verdict; prompt/config/input/output/result/cache/checkpoint/artifact hashes=required; evidence=2026-09-08; measurement versus assumption=explicit | Unavailable — unresolved routing confidence threshold has no matched, dated bilateral observation. | FAIL CLOSED — manual, probe, or source evidence required |
First-party provenance: Mistral AI model pricing; verification date 2026-09-08. Missing or conflicting joins fail closed.
Method and limitations: formulas are deterministic; observed and assumed inputs are labeled; no missing provider, host, account, region, realm, alias, snapshot, revision, weight, artifact, control, tool, modality, workload, rate-period, timestamp, or result is transferred. Run the mistral-small Batch 68 scenario →
How fast is Mistral Small 3.1?
How much does Mistral Small 3.1 cost at scale?
| Tokens / month | Est. cost (blended 3:1) |
|---|---|
| 100,000 | $0.03 |
| 1,000,000 | $0.26 |
| 10,000,000 | $2.62 |
| 100,000,000 | $26.25 |
How does Mistral Small 3.1 compare with other models?
What is Mistral Small 3.1 best for?
What should you explore next for Mistral Small 3.1?
What are common questions about Mistral Small 3.1?
Is Mistral Small 3.1 cheaper than GPT-4o Mini?
Mistral Small 3.1 costs $0.26/M blended tokens, GPT-4o Mini costs $0.26/M — GPT-4o Mini is cheaper.
How much does 1 million tokens cost with Mistral Small 3.1?
At a 3:1 input:output ratio, 1 million blended tokens costs approximately $0.26. Pure input costs $0.15/M; pure output costs $0.60/M.
What does Mistral Small 3.1 cost at high volume?
At 100 million blended tokens a month, Mistral Small 3.1 costs approximately $26.25. See the cost-at-scale table below for other volumes.
