Ministral 3 3B (Unsloth)
What most Nano models don’t offer: native multimodal input and tool calling right out of the box. Ministral 3 3B by Mistral AI delivers exactly that in the 3B class, with 256,000 tokens of context, an Apache 2.0 license, and local Unsloth GGUF distribution.
- Open Weights
- Nano
- llama.cpp
- Text
- Vision
- Instruction-Tuned
- Real-Time
Sovereign Risk: LOW TODO
Key metrics
Score · Latency · Cost · Quality
- Total Score Bronze
- 64.75
- Routine
- 39.27
- Reasoning
- 25.48
- LLM Judge Avg
- 3.14 / 5
- 100 Coverage
- Avg Task Duration
- 17.57s
- Real-Time
- Token Rate
- 53.04tok/s
- Output Rate
- P95 Latency
- 41.85s
- Top 5 %
- Total Tokens
- 93700
- Output Volume
- Cost per 1K
- $0
- USD / 1K Requests
- Benchmark Cost
- $0
- Total · 93700 tok
Benchmark modules
10 modules · weighted · vs. model median & top performer
Token efficiency & latency
Consumption per module vs. model median