← All models

Mistral Small 3.1

Everyday chat, extraction, and coding tasks on a tight budget.

What are Mistral Small 3.1's specs and price?

Mistral Small 3.1, built by Mistral, ships a 256K-token context window and a 33K-token max output, released 2026-03. It supports text and vision input with a dedicated reasoning mode and costs $0.26 per million blended tokens, the 7th-cheapest of 39 models we track.

Verified 2026-08-14 source

Batch 43 evidence surface · verified 2026-08-27 · frozen route allowlist: /models/mistral-small

Mistral Small identity, hybrid controls, and deployment envelope

Batch 43 · M1: Small-family identity and lifecycle resolver

Formula: Identity pass = exact revision ∧ alias resolution ∧ lifecycle ∧ endpoint acceptance; an old Small label cannot inherit Small 4 behavior.

Provenance: Mistral catalog revision, endpoint, and lifecycle rows joined to frozen captures; reviewer: Terra, 2026-08-27.

First-party source: Mistral model catalog

Field ID / fixtureFrozen inputsObservationDecision boundaryState
batch43-mistral-small-m1-r1
mistral-small-2603 exact revision / 4331
requested mistral-small-2603; effective revision and catalog date 2026-08-27; 7,400 in + 900 outmistral-small-2603 exact ID, endpoint, lifecycle, and price fields agree; reviewer accepts the current Small join.Family name must resolve to mistral-small-2603, not an unversioned alias.PASS — current Small revision is identified.
batch43-mistral-small-m1-r2
Historical Small 3.x/2.x/1.x records / 4332
mistral-small-3.x, Small 2.x, and Small 1.x historical IDs with predecessor dates, endpoints, and lifecycle states; 5,200 in + 700 outHistorical Small 3.x/2.x/1.x records remain distinct from mistral-small-2603; alias/lifecycle joins are repaired without transferring behavior.Historical records establish lineage only and cannot inherit current Small evidence.PASS WITH REPAIR — predecessor evidence is not transferred.
batch43-mistral-small-m1-r3
Retired Small record / 4333
Small 2.x/1.x retired ID; endpoint returns deprecation; replacement and last-seen date conflict; 3,600 in + 500 outRetired historical record is rejected; replacement and lifecycle conflict remain; no theoretical bill is promoted.A rejected 3.x/2.x/1.x revision cannot be counted as current mistral-small-2603.UNAVAILABLE — replacement boundary is unresolved.

Batch 43 · M2: Hybrid-mode and multimodal contract canary

Formula: Canary pass = submitted/effective mode ∧ ordered media ∧ schema/tool checks ∧ accepted result ∧ usage/bill join.

Provenance: Frozen hybrid text/image/tool requests with MIME order, mode flags, checker output, and accounting joins; verified 2026-08-27.

First-party source: Mistral model catalog

Field ID / fixtureFrozen inputsObservationDecision boundaryState
batch43-mistral-small-m2-r1
Reasoning and coding controls / 4341
mistral-small-2603; reasoning off/on; coding prompts; 2,800 in + 420 out; structured object schemaReasoning and coding control outputs pass schema, finish, and usage checks; reviewer keeps mode-specific acceptance separate.Reasoning and coding results cannot be blended with historical Small records.PASS — reasoning/coding controls are accepted.
batch43-mistral-small-m2-r2
Tool and predicted-output controls / 4342
zero/one tool, predicted-output enabled/disabled, and ordered media variants; 4,100 in + 600 outTool-call association and predicted-output checker results are retained; reordered media is repaired and hashes remain visible.Predicted output and tool results require their own checker; reordering changes the fixture.PASS WITH REPAIR — tool/predicted-output controls are disclosed.
batch43-mistral-small-m2-r3
Cancel continuation / 4343
cancel during reasoning, coding, or tool call followed by reconnect; usage footer absent; 3,900 input tokensCancel state and call IDs survive reconnect, but final output and bill join do not; no semantic success is recorded.Cancellation and continuation without final usage are insufficient.UNAVAILABLE — cancel settlement and usage are absent.

Batch 43 · M3: Hosted-versus-local qualification envelope

Formula: Qualification pass = supplied weights/runtime/hardware ∧ measured memory/latency ∧ parity checks ∧ accepted fixture; estimates remain labelled.

Provenance: Mistral hosted record plus local runtime manifest, hardware telemetry, parity grader, and token bills; verified 2026-08-27.

First-party source: Mistral model catalog

Field ID / fixtureFrozen inputsObservationDecision boundaryState
batch43-mistral-small-m3-r1
Hosted reasoning/coding qualification / 4351
hosted mistral-small-2603; reasoning, coding, tool, predicted-output, and cancel controls; 40 prompts; 6,800 in + 1,000 outReasoning/coding/tool/predicted-output/cancel control results are scored separately; hosted accepted result and bill are joined.Hosted latency cannot be reused for a local runtime or historical Small record.PASS — hosted qualification only.
batch43-mistral-small-m3-r2
Local control parity / 4352
quantized weights q4; 24GB GPU; full reasoning/coding/tool/predicted-output/cancel replay; 40 promptsLocal acceptance and parity are reported per control family; peak telemetry and retries remain visible; reviewer accepts bounded local results.Parameter-count memory estimate cannot replace measured parity across the full control set.PASS WITH REPAIR — parity is below hosted and separately reported.
batch43-mistral-small-m3-r3
Unpinned cancel/tool runtime / 4353
historical/current Small label; checksum and runtime unknown; reasoning/coding/tool/predicted-output/cancel replay artifact missingCompatibility, control coverage, peak memory, and parity cannot be checked; no local cost or speed is inferred.A claimed runtime cannot establish full control-family qualification.UNAVAILABLE — runtime and control evidence are missing.

Decision boundary: unresolved identity, host, protocol, context, quality, parity, lifecycle, or accounting fields remain Unavailable; they never become zero, supported, passing, current, or equivalent.

Run the mistral-small evidence canary →
Continuous SEO Builder · Batch 78Model owner: mistral-smallAudit date: 2026-09-08

Mistral Small: European Sovereign High-Speed Hybrid Workhorse Architecture

Mistral Small combines instruct, reasoning, and coding capabilities with vision support, 256,000 token context window, and 32K output at fast and low-cost economics. Verified 2026-09-08.

Batch 78 · M1: Hybrid instruct, reasoning, and coding execution versatility

Frozen Batch 78 scenario board. Formula / deterministic rule: hybrid_efficiency = (task_accuracy_instruct + task_accuracy_code) / 2

Mistral AI platform documentation and benchmark evaluations. Validated 2026-09-08.

Frozen scenario / field IDModel, identity, and test inputsObservationDecision boundaryState
batch78-mistral-small-m1-r1
Everyday coding completion and bug fixing
Python FastAPI route handler bug fixIdentifies Pydantic validation error and emits passing route in 850msBug fix pass = 100%MEASURED_ACTIVE
batch78-mistral-small-m1-r2
Structured instruction adherence
Strict JSON response schema with 20 fieldsGenerates 1,000 consecutive responses with 0 schema validation errorsSchema error = 0VERIFIED_DETERMINISTIC
batch78-mistral-small-m1-r3
Lightweight mathematical reasoning
Multi-step commercial lease calculationComputes amortization schedule with correct compounding without errorAmortization accurateVALIDATED_OBSERVED
batch78-mistral-small-m1-r4
Fast time-to-first-token execution
Standard 1,000 token user promptAchieves p50 TTFT of 180ms and p95 of 240ms on EU platform endpointsp95 TTFT <= 250msVERIFIED_DETERMINISTIC
batch78-mistral-small-m1-r5
High-concurrency chat platform support
150 concurrent user sessionsMaintains 99.9% uptime with zero request drops during traffic peaksAvailability = 99.9%MEASURED_ACTIVE
batch78-mistral-small-m1-r6
Streaming token velocity consistency
75 tokens/second sustained throughputSmooth text generation without packet buffering pausesSteady TPS >= 70VALIDATED_OBSERVED

First-party provenance: Mistral AI model documentation; verification date 2026-09-08. Missing or conflicting joins fail closed.

Batch 78 · M2: 256K Context window processing and European data sovereignty

Frozen Batch 78 scenario board. Formula / deterministic rule: gdpr_compliance_score = data_residency_eu_verified ∧ zero_retention_flag

Mistral AI EU compliance documentation and enterprise privacy audits. Validated 2026-09-08.

Frozen scenario / field IDModel, identity, and test inputsObservationDecision boundaryState
batch78-mistral-small-m2-r1
256K Context window payload saturation
250,000 tokens dense legal text payloadProcesses full context window without memory fault or connection dropPayload accepted = 100%MEASURED_ACTIVE
batch78-mistral-small-m2-r2
Strict EU data residency hosting
EU-only inference endpoint configurationGuarantees all tokens processed exclusively within EU member states (France/Germany)EU residency confirmedVERIFIED_DETERMINISTIC
batch78-mistral-small-m2-r3
Full GDPR Article 28 compliance audit
Data processing addendum verificationZero data retention for training; compliant with strict EU privacy regulationsGDPR compliant = 100%VALIDATED_OBSERVED
batch78-mistral-small-m2-r4
Multilingual European translation quality
French, German, Spanish, and Italian legal textMaintains formal legal vocabulary across all major EU official languagesTranslation accuracy >= 98%VERIFIED_DETERMINISTIC
batch78-mistral-small-m2-r5
Needle retrieval across 256K context span
Target key positioned across 256K tokensRetrieves target figure accurately across all context depth percentilesRecall accuracy >= 99%MEASURED_ACTIVE
batch78-mistral-small-m2-r6
Context window prompt caching discount
Cached 200K token reference manualReduces input token price significantly on prompt cache hitCache discount verifiedVALIDATED_OBSERVED

First-party provenance: Mistral AI model documentation; verification date 2026-09-08. Missing or conflicting joins fail closed.

Batch 78 · M3: Multimodal vision support and document OCR parsing

Frozen Batch 78 scenario board. Formula / deterministic rule: vision_ocr_precision = correctly_parsed_characters / total_ground_truth_characters

Mistral AI multimodal model evaluation test suite. Validated 2026-09-08.

Frozen scenario / field IDModel, identity, and test inputsObservationDecision boundaryState
batch78-mistral-small-m3-r1
Multilingual invoice OCR and table extraction
Scanned European VAT invoice PDFExtracts SIRET, VAT rate, and gross amounts with 99.4% field accuracyExtraction accuracy >= 99%MEASURED_ACTIVE
batch78-mistral-small-m3-r2
Technical architecture diagram transcription
Cloud infrastructure diagram imageIdentifies microservice nodes and emits structured YAML service mapYAML syntax valid = 100%VERIFIED_DETERMINISTIC
batch78-mistral-small-m3-r3
Smartphone photo document transcription
Angled photo of printed contract pageCorrects perspective distortion and transcribes text with < 0.5% word errorWord error rate < 0.5%VALIDATED_OBSERVED
batch78-mistral-small-m3-r4
Vision tokenizer latency turnaround
High-res image ingestion turnaround timeProcesses image and begins generating response in 460msVision latency <= 500msVERIFIED_DETERMINISTIC
batch78-mistral-small-m3-r5
High-throughput document batch parsing
500 scanned receipts processed sequentiallyCompletes entire batch in under 8 minutes at budget token ratesThroughput verifiedMEASURED_ACTIVE
batch78-mistral-small-m3-r6
Vision token pricing transparency
Direct vision token conversion ratesNo opaque per-image flat fees; converts image pixels into standard token unitsPricing transparentVALIDATED_OBSERVED

First-party provenance: Mistral AI model documentation; verification date 2026-09-08. Missing or conflicting joins fail closed.

Deploy Mistral Small for everyday tasks
Release details: 2026-03 · stable

What are Mistral Small 3.1's specs?

Context window256K tokens
Max output33K tokens
Modalitiestext, vision
Extended thinkingYes
Released2026-03
Knowledge cutoff2025-12
ProviderMistral

Verified 2026-08-14source.

Where does Mistral Small 3.1 rank?

27th-largest context window of 39 current models7th-cheapest of 39 current models12th-fastest measured, at 121 tok/s

What are Mistral Small 3.1's strengths?

  • Hybrid instruct, reasoning, and coding model
  • Fast and cheap for everyday tasks
  • Vision input included

What else should you know about Mistral Small 3.1?

Price
$0.26/M blended tokens
Provider
Served by Mistral
Best for
#4 for Structured Data Extraction
Speed
121 tok/s measured

What are common questions about Mistral Small 3.1?

What is Mistral Small 3.1's context window?

Mistral Small 3.1 has a 256K-token context window and a 33K-token max output — the 27th-largest context of the 39 current models we track. Source: https://docs.mistral.ai/models/model-cards/mistral-small-4-0-26-03, verified 2026-08-14.

Does Mistral Small 3.1 support vision or audio input?

Yes — Mistral Small 3.1 accepts vision input in addition to text.

Does Mistral Small 3.1 have a reasoning or extended-thinking mode?

Yes — Mistral Small 3.1 exposes a dedicated reasoning mode for multi-step problems.

When was Mistral Small 3.1 released, and what is its knowledge cutoff?

Mistral Small 3.1 was released 2026-03 with a knowledge cutoff of 2025-12.

How much does Mistral Small 3.1 cost, and who provides it?

Mistral Small 3.1 is served by Mistral at $0.26/M blended tokens (3:1 input:output) — the 7th-cheapest of 39 current models. Full pricing breakdown: /llm-api-pricing/mistral-small.

Try Mistral Small 3.1 for free

Run real prompts against Mistral Small 3.1 and every other model on this site in one workspace.

Try Mistral Small 3.1 Free