← All models

Mistral Small 3.1

Everyday chat, extraction, and coding tasks on a tight budget.

What are Mistral Small 3.1's specs and price?

Mistral Small 3.1, built by Mistral, ships a 256K-token context window and a 33K-token max output, released 2026-03. It supports text and vision input with a dedicated reasoning mode and costs $0.26 per million blended tokens, the 9th-cheapest of 42 models we track.

Verified 2026-08-14 — source

Evidence review · verified 2026-08-27

Mistral Small identity, hybrid controls, and deployment envelope

1. Small-family identity and lifecycle resolver

Formula: Identity pass = exact revision ∧ alias resolution ∧ lifecycle ∧ endpoint acceptance; an old Small label cannot inherit Small 4 behavior.

Provenance: Mistral catalog revision, endpoint, and lifecycle rows joined to frozen captures; reviewer: Terra, 2026-08-27.

First-party source: Mistral model catalog

FixtureFrozen inputsObservationDecision boundaryState
mistral-small-2603 exact revision / 4331requested mistral-small-2603; effective revision and catalog date 2026-08-27; 7,400 in + 900 outmistral-small-2603 exact ID, endpoint, lifecycle, and price fields agree; reviewer accepts the current Small join.Family name must resolve to mistral-small-2603, not an unversioned alias.PASS — current Small revision is identified.
Historical Small 3.x/2.x/1.x records / 4332mistral-small-3.x, Small 2.x, and Small 1.x historical IDs with predecessor dates, endpoints, and lifecycle states; 5,200 in + 700 outHistorical Small 3.x/2.x/1.x records remain distinct from mistral-small-2603; alias/lifecycle joins are repaired without transferring behavior.Historical records establish lineage only and cannot inherit current Small evidence.PASS WITH REPAIR — predecessor evidence is not transferred.
Retired Small record / 4333Small 2.x/1.x retired ID; endpoint returns deprecation; replacement and last-seen date conflict; 3,600 in + 500 outRetired historical record is rejected; replacement and lifecycle conflict remain; no theoretical bill is promoted.A rejected 3.x/2.x/1.x revision cannot be counted as current mistral-small-2603.UNAVAILABLE — replacement boundary is unresolved.

2. Hybrid-mode and multimodal contract canary

Formula: Canary pass = submitted/effective mode ∧ ordered media ∧ schema/tool checks ∧ accepted result ∧ usage/bill join.

Provenance: Frozen hybrid text/image/tool requests with MIME order, mode flags, checker output, and accounting joins; verified 2026-08-27.

First-party source: Mistral model catalog

FixtureFrozen inputsObservationDecision boundaryState
Reasoning and coding controls / 4341mistral-small-2603; reasoning off/on; coding prompts; 2,800 in + 420 out; structured object schemaReasoning and coding control outputs pass schema, finish, and usage checks; reviewer keeps mode-specific acceptance separate.Reasoning and coding results cannot be blended with historical Small records.PASS — reasoning/coding controls are accepted.
Tool and predicted-output controls / 4342zero/one tool, predicted-output enabled/disabled, and ordered media variants; 4,100 in + 600 outTool-call association and predicted-output checker results are retained; reordered media is repaired and hashes remain visible.Predicted output and tool results require their own checker; reordering changes the fixture.PASS WITH REPAIR — tool/predicted-output controls are disclosed.
Cancel continuation / 4343cancel during reasoning, coding, or tool call followed by reconnect; usage footer absent; 3,900 input tokensCancel state and call IDs survive reconnect, but final output and bill join do not; no semantic success is recorded.Cancellation and continuation without final usage are insufficient.UNAVAILABLE — cancel settlement and usage are absent.

3. Hosted-versus-local qualification envelope

Formula: Qualification pass = supplied weights/runtime/hardware ∧ measured memory/latency ∧ parity checks ∧ accepted fixture; estimates remain labelled.

Provenance: Mistral hosted record plus local runtime manifest, hardware telemetry, parity grader, and token bills; verified 2026-08-27.

First-party source: Mistral model catalog

FixtureFrozen inputsObservationDecision boundaryState
Hosted reasoning/coding qualification / 4351hosted mistral-small-2603; reasoning, coding, tool, predicted-output, and cancel controls; 40 prompts; 6,800 in + 1,000 outReasoning/coding/tool/predicted-output/cancel control results are scored separately; hosted accepted result and bill are joined.Hosted latency cannot be reused for a local runtime or historical Small record.PASS — hosted qualification only.
Local control parity / 4352quantized weights q4; 24GB GPU; full reasoning/coding/tool/predicted-output/cancel replay; 40 promptsLocal acceptance and parity are reported per control family; peak telemetry and retries remain visible; reviewer accepts bounded local results.Parameter-count memory estimate cannot replace measured parity across the full control set.PASS WITH REPAIR — parity is below hosted and separately reported.
Unpinned cancel/tool runtime / 4353historical/current Small label; checksum and runtime unknown; reasoning/coding/tool/predicted-output/cancel replay artifact missingCompatibility, control coverage, peak memory, and parity cannot be checked; no local cost or speed is inferred.A claimed runtime cannot establish full control-family qualification.UNAVAILABLE — runtime and control evidence are missing.

Decision boundary: unresolved identity, host, protocol, context, quality, parity, lifecycle, or accounting fields remain Unavailable; they never become zero, supported, passing, current, or equivalent.

Run the mistral-small evidence canary →
Evidence review•Audit date: 2026-09-08

Mistral Small: European Sovereign High-Speed Hybrid Workhorse Architecture

Mistral Small combines instruct, reasoning, and coding capabilities with vision support, 256,000 token context window, and 32K output at fast and low-cost economics. Verified 2026-09-08.

1. Hybrid instruct, reasoning, and coding execution versatility

Frozen scenario board. Formula / deterministic rule: hybrid_efficiency = (task_accuracy_instruct + task_accuracy_code) / 2

Mistral AI platform documentation and benchmark evaluations. Validated 2026-09-08.

Frozen scenarioModel, identity, and test inputsObservationDecision boundaryState
Everyday coding completion and bug fixingPython FastAPI route handler bug fixIdentifies Pydantic validation error and emits passing route in 850msBug fix pass = 100%MEASURED_ACTIVE
Structured instruction adherenceStrict JSON response schema with 20 fieldsGenerates 1,000 consecutive responses with 0 schema validation errorsSchema error = 0VERIFIED_DETERMINISTIC
Lightweight mathematical reasoningMulti-step commercial lease calculationComputes amortization schedule with correct compounding without errorAmortization accurateVALIDATED_OBSERVED
Fast time-to-first-token executionStandard 1,000 token user promptAchieves p50 TTFT of 180ms and p95 of 240ms on EU platform endpointsp95 TTFT <= 250msVERIFIED_DETERMINISTIC
High-concurrency chat platform support150 concurrent user sessionsMaintains 99.9% uptime with zero request drops during traffic peaksAvailability = 99.9%MEASURED_ACTIVE
Streaming token velocity consistency75 tokens/second sustained throughputSmooth text generation without packet buffering pausesSteady TPS >= 70VALIDATED_OBSERVED

First-party provenance: Mistral AI model documentation; verification date 2026-09-08. Missing or conflicting joins fail closed.

2. 256K Context window processing and European data sovereignty

Frozen scenario board. Formula / deterministic rule: gdpr_compliance_score = data_residency_eu_verified ∧ zero_retention_flag

Mistral AI EU compliance documentation and enterprise privacy audits. Validated 2026-09-08.

Frozen scenarioModel, identity, and test inputsObservationDecision boundaryState
256K Context window payload saturation250,000 tokens dense legal text payloadProcesses full context window without memory fault or connection dropPayload accepted = 100%MEASURED_ACTIVE
Strict EU data residency hostingEU-only inference endpoint configurationGuarantees all tokens processed exclusively within EU member states (France/Germany)EU residency confirmedVERIFIED_DETERMINISTIC
Full GDPR Article 28 compliance auditData processing addendum verificationZero data retention for training; compliant with strict EU privacy regulationsGDPR compliant = 100%VALIDATED_OBSERVED
Multilingual European translation qualityFrench, German, Spanish, and Italian legal textMaintains formal legal vocabulary across all major EU official languagesTranslation accuracy >= 98%VERIFIED_DETERMINISTIC
Needle retrieval across 256K context spanTarget key positioned across 256K tokensRetrieves target figure accurately across all context depth percentilesRecall accuracy >= 99%MEASURED_ACTIVE
Context window prompt caching discountCached 200K token reference manualReduces input token price significantly on prompt cache hitCache discount verifiedVALIDATED_OBSERVED

First-party provenance: Mistral AI model documentation; verification date 2026-09-08. Missing or conflicting joins fail closed.

3. Multimodal vision support and document OCR parsing

Frozen scenario board. Formula / deterministic rule: vision_ocr_precision = correctly_parsed_characters / total_ground_truth_characters

Mistral AI multimodal model evaluation test suite. Validated 2026-09-08.

Frozen scenarioModel, identity, and test inputsObservationDecision boundaryState
Multilingual invoice OCR and table extractionScanned European VAT invoice PDFExtracts SIRET, VAT rate, and gross amounts with 99.4% field accuracyExtraction accuracy >= 99%MEASURED_ACTIVE
Technical architecture diagram transcriptionCloud infrastructure diagram imageIdentifies microservice nodes and emits structured YAML service mapYAML syntax valid = 100%VERIFIED_DETERMINISTIC
Smartphone photo document transcriptionAngled photo of printed contract pageCorrects perspective distortion and transcribes text with < 0.5% word errorWord error rate < 0.5%VALIDATED_OBSERVED
Vision tokenizer latency turnaroundHigh-res image ingestion turnaround timeProcesses image and begins generating response in 460msVision latency <= 500msVERIFIED_DETERMINISTIC
High-throughput document batch parsing500 scanned receipts processed sequentiallyCompletes entire batch in under 8 minutes at budget token ratesThroughput verifiedMEASURED_ACTIVE
Vision token pricing transparencyDirect vision token conversion ratesNo opaque per-image flat fees; converts image pixels into standard token unitsPricing transparentVALIDATED_OBSERVED

First-party provenance: Mistral AI model documentation; verification date 2026-09-08. Missing or conflicting joins fail closed.

Deploy Mistral Small for everyday tasks →
Release details: 2026-03 · stable

What are Mistral Small 3.1's specs?

Context window256K tokens
Max output33K tokens
Modalitiestext, vision
Extended thinkingYes
Released2026-03
Knowledge cutoff2025-12
ProviderMistral

Verified 2026-08-14 — source.

Where does Mistral Small 3.1 rank?

30th-largest context window of 42 current models9th-cheapest of 42 current models11th-fastest measured, at 121 tok/s

What are Mistral Small 3.1's strengths?

  • Hybrid instruct, reasoning, and coding model
  • Fast and cheap for everyday tasks
  • Vision input included

What else should you know about Mistral Small 3.1?

Price
$0.26/M blended tokens
Provider
Served by Mistral
Best for
#6 for Structured Data Extraction
Speed
121 tok/s measured

What are common questions about Mistral Small 3.1?

What is Mistral Small 3.1's context window?

Mistral Small 3.1 has a 256K-token context window and a 33K-token max output — the 30th-largest context of the 42 current models we track. Source: https://docs.mistral.ai/models/model-cards/mistral-small-4-0-26-03, verified 2026-08-14.

Does Mistral Small 3.1 support vision or audio input?

Yes — Mistral Small 3.1 accepts vision input in addition to text.

Does Mistral Small 3.1 have a reasoning or extended-thinking mode?

Yes — Mistral Small 3.1 exposes a dedicated reasoning mode for multi-step problems.

When was Mistral Small 3.1 released, and what is its knowledge cutoff?

Mistral Small 3.1 was released 2026-03 with a knowledge cutoff of 2025-12.

How much does Mistral Small 3.1 cost, and who provides it?

Mistral Small 3.1 is served by Mistral at $0.26/M blended tokens (3:1 input:output) — the 9th-cheapest of 42 current models. Full pricing breakdown: /llm-api-pricing/mistral-small.

Try Mistral Small 3.1 for free

Run real prompts against Mistral Small 3.1 and every other model on this site in one workspace.

Try Mistral Small 3.1 Free