← All open-weight models

Mistral AI · Vision & multimodal

Mistral Small 4

Best price/performance workhorse — a hybrid instruct+reasoning+coding model with only 6.5B active params, so it is fast and cheap while still handling agents and vision; the model to self-host under Apache 2.0 if you want reasoning without a licence conversation. Alias: mistral-small-latest.

Open weights · Permissive (Apache, MIT, OpenMDW) · Data checked September 6, 2026 · Not a hands-on evaluation

Model ID mistral-small-2603

Parameters
119B total · 6.5B active119B-A6.5B MoE
Architecture
Mixture of experts
Context window
256,000 tokensMaximum input
Max output
Not recordedPer response
Input price
$0.15 per 1M tokens
Output price
$0.60 per 1M tokens
Cached input
$0.015 per 1M tokens
Blended price
$0.2625 per 1M tokens3:1 input to output
Knowledge cutoff
Not recorded
Released
2026-03-16
Status
ga
Size band
40B to 120B

License and openness

Mistral Small 4 is released under the Apache-2.0. This is a permissive license. It generally allows use, modification and redistribution, including commercial use, subject to attribution and notice requirements. License terms can change between versions; confirm the text published with the weights you download.

Capabilities

Input modalitiestext, image, pdf
Output modalitiestext

What it is for

Alias: mistral-small-latest.

At a three-to-one input-to-output ratio, Mistral Small 4 costs $0.26 per million tokens blended. A workload of one million input and 330,000 output tokens per day would run about $10.44 per month at list price, before caching or batch discounts.

Cost at list price

Input$0.15 per 1M tokens
Output$0.60 per 1M tokens
Cached input$0.015 per 1M tokens
Blended, 3:1$0.2625 per 1M tokens
Estimate a workload across all models ↗

Sources

Sources checked September 6, 2026. Specifications and prices change without notice; confirm against the provider before you commit.

Other open-weight models from Mistral AI

Browse all open-weight models ↗