Mistral

Mistral Small 3.1 (25.03)

Mistral Small 3.1 (25.03) is a text generation model developed by Mistral, released in March 2025. It features a 128,000-token context window, multimodal understanding, and support for dozens of spoken languages alongside more than 80 coding languages. The model is designed to run on a single node, making it practical for deployment without distributed infrastructure. This version introduces improved text performance and expanded context handling compared to earlier Mistral Small releases. At an inference speed of approximately 150 tokens per second, it is suited for tasks that require both throughput and long-context processing, such as document analysis, multilingual applications, and code generation. Its combination of broad language coverage and single-node efficiency makes it a practical choice for developers building production applications with constrained compute budgets.

Unknown 128,000 context 16,000 tokens output
Long Context Window Multilingual Text Code Generation Multimodal Understanding Fast Inference Function Calling

Model Overview

High-signal model metadata in a structured two-column overview table.

Provider

The entity that provides this model.

Mistral

Input Context Window

The number of tokens supported by the input context window.

128,000 tokens

Maximum Output Tokens

The number of tokens that can be generated by the model in a single request.

16,000 tokens tokens

Open Source

Whether the model's code is available for public use.

No

Release Date

When the model was first released.

Unknown

Knowledge Cut-off Date

When the model's knowledge was last updated.

Unknown

API Providers

The providers that offer this model. This is not an exhaustive list.

Mistral API, Hugging Face

Modalities

Types of data this model can process.

Text Code

What is Mistral Small 3.1 (25.03)

A fuller summary of positioning, capabilities, and source-specific details for Mistral Small 3.1 (25.03).

Mistral Small 3.1 (25.03) is a text generation model developed by Mistral, released in March 2025. It features a 128,000-token context window, multimodal understanding, and support for dozens of spoken languages alongside more than 80 coding languages. The model is designed to run on a single node, making it practical for deployment without distributed infrastructure.

This version introduces improved text performance and expanded context handling compared to earlier Mistral Small releases. At an inference speed of approximately 150 tokens per second, it is suited for tasks that require both throughput and long-context processing, such as document analysis, multilingual applications, and code generation. Its combination of broad language coverage and single-node efficiency makes it a practical choice for developers building production applications with constrained compute budgets.

Capabilities

What Mistral Small 3.1 (25.03) supports

CTX

Long Context Window

Processes up to 128,000 tokens in a single request, enabling analysis of long documents, codebases, or extended conversations without truncation.

AI

Multilingual Text

Supports dozens of spoken languages for generation and comprehension tasks, making it suitable for international and localized applications.

</>

Code Generation

Handles code tasks across 80+ programming languages, including generation, completion, and explanation.

MM

Multimodal Understanding

Accepts image inputs alongside text, allowing the model to reason about visual content within a single prompt.

AI

Fast Inference

Delivers approximately 150 tokens per second, supporting latency-sensitive production workloads on a single node.

AI

Function Calling

Supports structured tool use and function calling, enabling integration with external APIs and agentic workflows.

Pricing for Mistral Small 3.1 (25.03)

Primary API pricing shown in the same “quick compare” spirit as the reference page.

Price Comparison

Additional usage-cost dimensions synced into the project for this model.

maxTemperature 1
maxResponseSize 16,000 tokens

API Access & Providers

Places where this model is available, based on the synced detail-page metadata.

Mistral API Hugging Face

Model Performance

Benchmark scores synced from the current model source and normalized into the local catalog.

Benchmark Score
AIME 2024
American math olympiad problems
6.3%
GPQA Diamond
PhD-level science questions (biology, physics, chemistry)
38.1%
HLE
Questions that challenge frontier models across many domains
4.3%
LiveCodeBench
Real-world coding tasks from recent competitions
14.1%
MATH-500
Undergraduate and competition-level math problems
56.3%
MMLU-Pro
Expert knowledge across 14 academic disciplines
52.9%
SciCode
Scientific research coding and numerical methods
15.6%

Resources & Documentation

Official model cards, release notes, docs, and other references synced from the source page.

Compare Mistral Small 3.1 (25.03) with related models

Jump straight into the most relevant side-by-side comparison pages for this model.

Related Daily Briefs

Recent daily stories tied to Mistral Small 3.1 (25.03) through direct model mentions or provider-level coverage.

FAQ

Common questions about Mistral Small 3.1 (25.03)

What is the context window size for Mistral Small 3.1 (25.03)?

The model supports a context window of 128,000 tokens, allowing it to process long documents or extended conversations in a single request.

Does Mistral Small 3.1 (25.03) support image inputs?

Yes. This version includes multimodal understanding, meaning it can accept and reason about image inputs in addition to text.

How many coding languages does this model support?

The model supports over 80 coding languages, making it broadly applicable for code generation, completion, and explanation tasks.

What is the knowledge cutoff date for this model?

A specific training data cutoff date is not listed in the available metadata for this model version.

Can this model run on a single machine?

Yes. Mistral Small 3.1 (25.03) is designed for single-node inference, meaning it does not require distributed compute infrastructure to run.

More models from Mistral

Continue browsing adjacent models from the same provider.

← All AI Models