Ministral 3 3B (Unsloth)

What most Nano models don’t offer: native multimodal input and tool calling right out of the box. Ministral 3 3B by Mistral AI delivers exactly that in the 3B class, with 256,000 tokens of context, an Apache 2.0 license, and local Unsloth GGUF distribution.

Mistral AI Version 3 Commercial use permitted Dense 3 B (3 B active) 256 K Context 07/2025 locally tested

  • Open Weights
  • Nano
  • llama.cpp
  • Text
  • Vision
  • Instruction-Tuned
  • Real-Time

Sovereign Risk: LOW TODO

Key metrics

Score · Latency · Cost · Quality

Total Score Bronze
64.75
Routine
39.27
Reasoning
25.48

Rank #88

LLM Judge Avg
3.14
100 Coverage
Avg Task Duration
17.57
Real-Time
Token Rate
53.04
Output Rate
P95 Latency
41.85
Top 5 %
Total Tokens
93700
Output Volume
Cost per 1K
$0
USD / 1K Requests
Benchmark Cost
$0
Total · 93700 tok

Benchmark modules

10 modules · weighted · vs. model median & top performer

Ministral 3 3B (Unsloth) Best model Ø All models
Code Quality 65.3
CLI Benchmark 75
Logical Reasoning 62.09
UX Writing 67.99
Documentation 68.23
Content Transform. 71.77
Cultural Intelligence 56.3
Synthesis Quality 32.5
Tool Execution 83.33
ToolUse Score 56.42
Benchmark Cost $0

Token efficiency & latency

Consumption per module vs. model median

Token consumption per module

Performance profile