Mistral API Cost Calculator

How much does the Mistral API cost per month?

For Ministral 8B at 200K requests/month, 2,400 input tokens and 350 output tokens per request costs about $40.88 per month after 30% cache use and 100% batch share. Across Mistral's 5 priced models, the cheapest default ranking is Ministral 8B at $81.77 per month.

Verified 2026-06-21

Pricing data as of June 2026. Sources: Mistral pricing and model documentation. Parameters are shareable in this URL.

Ministral 8B: estimated monthly cost

$40.88

Formula: provider-specific cached input multiplier × cache rate, plus uncached input and verbosity-adjusted output; cache writes are amortised over the provider TTL, then batch savings are applied.

Per request$0.000
Per day$1.36
Per month$40.88
Per year$491

OpenAI scenario sensitivity

The default is a server-rendered estimate. Change the cache and batch shares to see when a cheaper qualified model overtakes the selected model; the URL is shareable and preserves the inputs.

Default$40.88/month
No caching or batch$81.77/month
50% cache, 50% batch$61.32/month
100% cache, 100% batch$40.88/month

Batch 11 Mistral workload decision depth

1. Mixed synchronous/batch bill

Async shareMonthly billSavingsEligibility/turnaround
0%$82.50$0.0000Synchronous
25%$72.19$10.31Submitted eligible job; asynchronous deadline
50%$61.88$20.63Submitted eligible job; asynchronous deadline
100%$41.25$41.25Submitted eligible job; asynchronous deadline

2. Context and modality fit

Context inputModelToken billFit/modality boundary
32KMinistral 8B$0.0049Window and non-text units: sourced per model or Unavailable
32KMistral Small 3.1$0.0050Window and non-text units: sourced per model or Unavailable
32KCodestral$0.0099Window and non-text units: sourced per model or Unavailable
32KMistral Large 3$0.02Window and non-text units: sourced per model or Unavailable
128KMinistral 8B$0.02Window and non-text units: sourced per model or Unavailable
128KMistral Small 3.1$0.02Window and non-text units: sourced per model or Unavailable
128KCodestral$0.04Window and non-text units: sourced per model or Unavailable
128KMistral Large 3$0.06Window and non-text units: sourced per model or Unavailable
256KMinistral 8B$0.04Window and non-text units: sourced per model or Unavailable
256KMistral Small 3.1$0.04Window and non-text units: sourced per model or Unavailable
256KCodestral$0.08Window and non-text units: sourced per model or Unavailable
256KMistral Large 3$0.13Window and non-text units: sourced per model or Unavailable

3. Large/Medium/Small accepted-result crossover

TierAPI cost/callRequired upliftRetry inputAccepted formula
Ministral 8B$0.0004User-supplied quality upliftUser-supplied retry rateAPI cost ÷ (1 − retry rate)
Mistral Small 3.1$0.0006User-supplied quality upliftUser-supplied retry rateAPI cost ÷ (1 − retry rate)
Codestral$0.0010User-supplied quality upliftUser-supplied retry rateAPI cost ÷ (1 − retry rate)

Provenance: Batch 11 Mistral delivery/context/accepted-result module; selected model Ministral 8B; inputs are 2,400 input tokens, 350 output tokens, 200,000 calls/month, cache rate 30%, batch share 100%. Verified 2026-06-21. Provider source: https://mistral.ai/pricing · Run this scenario →

Batch 59 · server-rendered evidence boards · verified 2026-09-07

Intent answer: Mistral API billing spans cost-efficient edge models (Ministral 8B) through enterprise flagships (Mistral Large). Batch processing discounts, prompt caching support, and specialized Codestral pricing dictate optimum architecture. Verified 2026-09-07.

Demand evidence: Qualitative demand: Mistral pricing and token calculators reviewed 2026-09-07; exact US monthly volume is unavailable.

Scope boundary: Calculate Mistral API monthly expenses across Mistral Large, Mistral Medium, Mistral Small, Ministral 8B, and Codestral with prompt caching. Exact joins required; unresolved joins render Unavailable.

Mistral tier tiering & workload routing matrix

Deterministic formula / rule: tier_cost = min(input * rate_model + output * rate_model) subject to task complexity threshold.

Boundary: Owns Mistral portfolio routing economics.

Frozen scenario / field IDExact identity and evidence fieldsResultState
batch59-calc-mistral-m1-r1
simple classification (Ministral 8B)
route=$/llm-cost-calculator/mistral; owner=$calc-mistral; scenario=$simple classification (Ministral 8B); task token shape; candidate models; unit token costs; monthly bill per model; tier cost delta; selection threshold; verified=2026-09-07Unavailable — frozen calc-mistral fixture requires an exact source, identity, and result receiptUNTESTED — module 1 rule is reproducible but no production observation is claimed
batch59-calc-mistral-m1-r2
general reasoning (Mistral Small)
route=$/llm-cost-calculator/mistral; owner=$calc-mistral; scenario=$general reasoning (Mistral Small); task token shape; candidate models; unit token costs; monthly bill per model; tier cost delta; selection threshold; verified=2026-09-07Unavailable — frozen calc-mistral fixture requires an exact source, identity, and result receiptUNTESTED — module 1 rule is reproducible but no production observation is claimed
batch59-calc-mistral-m1-r3
complex reasoning (Mistral Large)
route=$/llm-cost-calculator/mistral; owner=$calc-mistral; scenario=$complex reasoning (Mistral Large); task token shape; candidate models; unit token costs; monthly bill per model; tier cost delta; selection threshold; verified=2026-09-07Unavailable — frozen calc-mistral fixture requires an exact source, identity, and result receiptUNTESTED — module 1 rule is reproducible but no production observation is claimed
batch59-calc-mistral-m1-r4
code completion (Codestral)
route=$/llm-cost-calculator/mistral; owner=$calc-mistral; scenario=$code completion (Codestral); task token shape; candidate models; unit token costs; monthly bill per model; tier cost delta; selection threshold; verified=2026-09-07Unavailable — frozen calc-mistral fixture requires an exact source, identity, and result receiptUNTESTED — module 1 rule is reproducible but no production observation is claimed
batch59-calc-mistral-m1-r5
multilingual document translation
route=$/llm-cost-calculator/mistral; owner=$calc-mistral; scenario=$multilingual document translation; task token shape; candidate models; unit token costs; monthly bill per model; tier cost delta; selection threshold; verified=2026-09-07Unavailable — frozen calc-mistral fixture requires an exact source, identity, and result receiptUNTESTED — module 1 rule is reproducible but no production observation is claimed
batch59-calc-mistral-m1-r6
unsupported model alias
route=$/llm-cost-calculator/mistral; owner=$calc-mistral; scenario=$unsupported model alias; task token shape; candidate models; unit token costs; monthly bill per model; tier cost delta; selection threshold; verified=2026-09-07Unavailable — exact calc-mistral evidence join is not closed for "unsupported model alias"FAIL CLOSED — manual, probe, or source evidence required

First-party citation: Mistral AI official pricing. Verified 2026-09-07; missing or conflicting joins fail closed.

Codestral developer team cost projection

Deterministic formula / rule: team_cost = developer_count * daily_requests * (avg_input * rate_in + avg_output * rate_out) * 22_workdays.

Boundary: Owns developer coding assistant budget forecasting.

Frozen scenario / field IDExact identity and evidence fieldsResultState
batch59-calc-mistral-m2-r1
small team (5 devs, 500 req/day)
route=$/llm-cost-calculator/mistral; owner=$calc-mistral; scenario=$small team (5 devs, 500 req/day); dev count; daily queries; prompt context size; completion size; monthly API total; per-seat equivalent cost; verified=2026-09-07Unavailable — frozen calc-mistral fixture requires an exact source, identity, and result receiptUNTESTED — module 2 rule is reproducible but no production observation is claimed
batch59-calc-mistral-m2-r2
medium team (25 devs, 2.5K req/day)
route=$/llm-cost-calculator/mistral; owner=$calc-mistral; scenario=$medium team (25 devs, 2.5K req/day); dev count; daily queries; prompt context size; completion size; monthly API total; per-seat equivalent cost; verified=2026-09-07Unavailable — frozen calc-mistral fixture requires an exact source, identity, and result receiptUNTESTED — module 2 rule is reproducible but no production observation is claimed
batch59-calc-mistral-m2-r3
large team (100 devs, 10K req/day)
route=$/llm-cost-calculator/mistral; owner=$calc-mistral; scenario=$large team (100 devs, 10K req/day); dev count; daily queries; prompt context size; completion size; monthly API total; per-seat equivalent cost; verified=2026-09-07Unavailable — frozen calc-mistral fixture requires an exact source, identity, and result receiptUNTESTED — module 2 rule is reproducible but no production observation is claimed
batch59-calc-mistral-m2-r4
automated CI code reviewer
route=$/llm-cost-calculator/mistral; owner=$calc-mistral; scenario=$automated CI code reviewer; dev count; daily queries; prompt context size; completion size; monthly API total; per-seat equivalent cost; verified=2026-09-07Unavailable — frozen calc-mistral fixture requires an exact source, identity, and result receiptUNTESTED — module 2 rule is reproducible but no production observation is claimed
batch59-calc-mistral-m2-r5
batch repository analysis
route=$/llm-cost-calculator/mistral; owner=$calc-mistral; scenario=$batch repository analysis; dev count; daily queries; prompt context size; completion size; monthly API total; per-seat equivalent cost; verified=2026-09-07Unavailable — frozen calc-mistral fixture requires an exact source, identity, and result receiptUNTESTED — module 2 rule is reproducible but no production observation is claimed
batch59-calc-mistral-m2-r6
unsupported IDE plugin rate
route=$/llm-cost-calculator/mistral; owner=$calc-mistral; scenario=$unsupported IDE plugin rate; dev count; daily queries; prompt context size; completion size; monthly API total; per-seat equivalent cost; verified=2026-09-07Unavailable — exact calc-mistral evidence join is not closed for "unsupported IDE plugin rate"FAIL CLOSED — manual, probe, or source evidence required

First-party citation: Mistral AI official pricing. Verified 2026-09-07; missing or conflicting joins fail closed.

Mistral prompt caching & batch savings board

Deterministic formula / rule: net_spend = (uncached_inputs * in_rate) + (cached_inputs * cache_rate) + (outputs * out_rate) - batch_discount.

Boundary: Owns Mistral optimization lever modeling.

Frozen scenario / field IDExact identity and evidence fieldsResultState
batch59-calc-mistral-m3-r1
50% cache hit enterprise workload
route=$/llm-cost-calculator/mistral; owner=$calc-mistral; scenario=$50% cache hit enterprise workload; monthly requests; cache hit ratio; standard cost; optimized cost; absolute monthly savings; ROI horizon; verified=2026-09-07Unavailable — frozen calc-mistral fixture requires an exact source, identity, and result receiptUNTESTED — module 3 rule is reproducible but no production observation is claimed
batch59-calc-mistral-m3-r2
80% cache hit RAG system
route=$/llm-cost-calculator/mistral; owner=$calc-mistral; scenario=$80% cache hit RAG system; monthly requests; cache hit ratio; standard cost; optimized cost; absolute monthly savings; ROI horizon; verified=2026-09-07Unavailable — frozen calc-mistral fixture requires an exact source, identity, and result receiptUNTESTED — module 3 rule is reproducible but no production observation is claimed
batch59-calc-mistral-m3-r3
asynchronous batch extraction pipeline
route=$/llm-cost-calculator/mistral; owner=$calc-mistral; scenario=$asynchronous batch extraction pipeline; monthly requests; cache hit ratio; standard cost; optimized cost; absolute monthly savings; ROI horizon; verified=2026-09-07Unavailable — frozen calc-mistral fixture requires an exact source, identity, and result receiptUNTESTED — module 3 rule is reproducible but no production observation is claimed
batch59-calc-mistral-m3-r4
streaming interactive zero-cache chat
route=$/llm-cost-calculator/mistral; owner=$calc-mistral; scenario=$streaming interactive zero-cache chat; monthly requests; cache hit ratio; standard cost; optimized cost; absolute monthly savings; ROI horizon; verified=2026-09-07Unavailable — frozen calc-mistral fixture requires an exact source, identity, and result receiptUNTESTED — module 3 rule is reproducible but no production observation is claimed
batch59-calc-mistral-m3-r5
mixed latency-tolerant job queue
route=$/llm-cost-calculator/mistral; owner=$calc-mistral; scenario=$mixed latency-tolerant job queue; monthly requests; cache hit ratio; standard cost; optimized cost; absolute monthly savings; ROI horizon; verified=2026-09-07Unavailable — frozen calc-mistral fixture requires an exact source, identity, and result receiptUNTESTED — module 3 rule is reproducible but no production observation is claimed
batch59-calc-mistral-m3-r6
unsupported cache TTL parameter
route=$/llm-cost-calculator/mistral; owner=$calc-mistral; scenario=$unsupported cache TTL parameter; monthly requests; cache hit ratio; standard cost; optimized cost; absolute monthly savings; ROI horizon; verified=2026-09-07Unavailable — exact calc-mistral evidence join is not closed for "unsupported cache TTL parameter"FAIL CLOSED — manual, probe, or source evidence required

First-party citation: Mistral AI official pricing. Verified 2026-09-07; missing or conflicting joins fail closed.

Method and limitations: this board exposes deterministic rules, first-party citations, and dated evidence identities. It does not invent volume, coverage, entitlement, retention, residency, quota, capacity, feature support, quality, price, reliability, or legal conclusions. Run the calc-mistral evidence flow →

Ranked cost — 200K requests/month

ModelProviderList monthlyEffective monthlyVerbosityRank Δ
Ministral 8BbudgetMistral$82.50$40.880.93×
Mistral Small 3.1budgetMistral$114$53.850.85×
CodestralbudgetMistral$207$96.890.79×
Mistral Large 3budgetMistral$345$173
Mistral Medium 3midMistral$1,245$5800.84×

Levers live on Mistral

Batch API: up to 50%Model verbosity: up to 95%Context trimming: up to 47%

The billing gotcha

Mistral's batch discount applies per submitted job, not per call, so it only pays off once you're already batching requests server-side rather than calling the synchronous API in a loop — turning batch on in the estimator without changing your actual call pattern will overstate your real saving, since a request sent one at a time never qualifies for the batch rate no matter what this page's toggle shows.

Related

Mistral provider profile →Compare against another provider →Global cost calculator →How to reduce LLM API costs →