Mistral Small 4
Mistral Small 4 is Mistral AI’s compact Open Weights model for general and agentic tasks. The MoE architecture activates only 6.5 billion of the total 119 billion parameters per token, the context window spans 256,000 tokens, and the model processes text and image inputs. Available under the Apache 2.0 license for local use or via the Mistral API, from a European provider environment.
- Open Weights
- Workstation
- Mistral AI
- Text
- Vision
- Instruction-Tuned
- Long Context
- Real-Time
Key metrics
Score · Latency · Cost · Quality
- Total Score Silver
- 73.3
- Routine
- 45.77
- Reasoning
- 27.54
- LLM Judge Avg
- 3.63 / 5
- 100 Coverage
- Avg Task Duration
- 6.22s
- Real-Time
- Token Rate
- 116.82tok/s
- Output Rate
- P95 Latency
- 18.63s
- Top 5 %
- Total Tokens
- 55000
- Output Volume
- Cost per 1K
- $0.0003
- USD / 1K Requests
- Benchmark Cost
- $0.02
- Total · 55000 tok
Benchmark modules
10 modules · weighted · vs. model median & top performer
Token efficiency & latency
Consumption per module vs. model median