Ministral 3 14B (Unsloth)

Ministral 3 14B ist das Top-Modell der Ministral-3-Familie für lokale Assistenten- und Analyse-Workloads. 13,9 Mrd. Dense-Parameter, 256.000 Tokens Kontext, multimodale Eingabe für Text und Bild, nativem Function-Calling und JSON-Output. Apache-2.0-lizenziert und als Unsloth-GGUF-Variante lokal voll betreibbar — der bisher leistungsfähigste Ministral ohne Cloud-Zwang.

Mistral AI Version 3 Kommerzielle Nutzung erlaubt Dense 13.9 B (13.9 B aktiv) 256 K Context 07/2025 local getestet

  • Open Weights
  • Desktop
  • llama.cpp
  • Text
  • Vision
  • Instruction-Tuned
  • Batch

Sovereign Risk: LOW TODO

Schlüsselmetriken

Score · Latenz · Kosten · Qualität

Total Score Silver
74.49
Routine
45.77
Reasoning
28.72

Rank #31

LLM Judge Avg
3.76
100 Coverage
Avg Task Duration
68.03
Batch
Token Rate
15.78
Output Rate
P95 Latency
162.1
Top 5 %
Total Tokens
100100
Output Volume
Cost per 1K
$0
USD / 1K Requests
Benchmark Cost
$0
Total · 100100 tok

Benchmark-Module

10 Module · gewichtet · vs. Modellmedian & Spitzenreiter

Ministral 3 14B (Unsloth) Bestes Modell Ø Alle Modelle
Code Quality 77.1
CLI Benchmark 81.12
Logical Reasoning 70.25
UX Writing 74.05
Documentation 76.63
Content Transform. 77.81
Cultural Intelligence 81.1
Synthesis Quality 33.33
Tool Execution 85.83
ToolUse Score 61.17
Benchmark Cost $0

Token-Effizienz & Latenz

Verbrauch pro Modul vs. Modellmedian

Token-Verbrauch pro Modul

Performance-Profil