Mistral Small 4

Mistral Small 4 is Mistral AI’s compact Open Weights model for general and agentic tasks. The MoE architecture activates only 6.5 billion of the total 119 billion parameters per token, the context window spans 256,000 tokens, and the model processes text and image inputs. Available under the Apache 2.0 license for local use or via the Mistral API, from a European provider environment.

Mistral AI Version 4 Commercial use permitted MoE 119 B (6.5 B active) 256 K Context 01/2026 $0.1 / $0.3 per 1M

  • Open Weights
  • Workstation
  • Mistral AI
  • Text
  • Vision
  • Instruction-Tuned
  • Long Context
  • Real-Time

Key metrics

Score · Latency · Cost · Quality

Total Score Silver
73.3
Routine
45.77
Reasoning
27.54

Rank #38

LLM Judge Avg
3.63
100 Coverage
Avg Task Duration
6.22
Real-Time
Token Rate
116.82
Output Rate
P95 Latency
18.63
Top 5 %
Total Tokens
55000
Output Volume
Cost per 1K
$0.0003
USD / 1K Requests
Benchmark Cost
$0.02
Total · 55000 tok

Benchmark modules

10 modules · weighted · vs. model median & top performer

Mistral Small 4 Best model Ø All models
Code Quality 70.92
CLI Benchmark 79.67
Logical Reasoning 72.47
UX Writing 72.35
Documentation 76.4
Content Transform. 72.61
Cultural Intelligence 71.44
Synthesis Quality 39.17
Tool Execution 66.67
ToolUse Score 20.33
Benchmark Cost $0.02

Token efficiency & latency

Consumption per module vs. model median

Token consumption per module

Performance profile