Ministral 3 14B (Unsloth)

Ministral 3 14B is the top model of the Ministral 3 family for local assistant and analysis workloads. 13.9B dense parameters, 256,000 tokens of context, multimodal input for text and image, native function calling and JSON output. Licensed under Apache 2.0 and fully operable locally as an Unsloth GGUF variant — the most capable Ministral to date without any cloud dependency.

Mistral AI Version 3 Commercial use permitted Dense 13.9 B (13.9 B active) 256 K Context 07/2025 locally tested

  • Open Weights
  • Desktop
  • llama.cpp
  • Text
  • Vision
  • Instruction-Tuned
  • Batch

Sovereign Risk: LOW TODO

Key metrics

Score · Latency · Cost · Quality

Total Score Silver
74.49
Routine
45.77
Reasoning
28.72

Rank #31

LLM Judge Avg
3.76
100 Coverage
Avg Task Duration
68.03
Batch
Token Rate
15.78
Output Rate
P95 Latency
162.1
Top 5 %
Total Tokens
100100
Output Volume
Cost per 1K
$0
USD / 1K Requests
Benchmark Cost
$0
Total · 100100 tok

Benchmark modules

10 modules · weighted · vs. model median & top performer

Ministral 3 14B (Unsloth) Best model Ø All models
Code Quality 77.1
CLI Benchmark 81.12
Logical Reasoning 70.25
UX Writing 74.05
Documentation 76.63
Content Transform. 77.81
Cultural Intelligence 81.1
Synthesis Quality 33.33
Tool Execution 85.83
ToolUse Score 61.17
Benchmark Cost $0

Token efficiency & latency

Consumption per module vs. model median

Token consumption per module

Performance profile