Phi-4 Mini (Unsloth)
3.8B dense parameters, synthetic training data, and a pronounced reasoning specialization: Phi-4-mini is Microsoft’s compact model for math, logic, and structured output. MIT license, 128,000 tokens of context, locally deployable as an Unsloth GGUF — not a broad generalist, but a specialist at Nano scale.
- Open Weights
- Nano
- llama.cpp
- Text
- Instruction-Tuned
- Real-Time
Sovereign Risk: LOW TODO
Key metrics
Score · Latency · Cost · Quality
- Total Score Bronze
- 58.41
- Routine
- 34.53
- Reasoning
- 23.88
- LLM Judge Avg
- 2.79 / 5
- 100 Coverage
- Avg Task Duration
- 17.44s
- Real-Time
- Token Rate
- 46.13tok/s
- Output Rate
- P95 Latency
- 20.51s
- Top 5 %
- Total Tokens
- 62900
- Output Volume
- Cost per 1K
- $0
- USD / 1K Requests
- Benchmark Cost
- $0
- Total · 62900 tok
Benchmark modules
10 modules · weighted · vs. model median & top performer
Token efficiency & latency
Consumption per module vs. model median