DeepSeek R1 Distill Qwen 14B
The strongest variant in the DeepSeek-R1-Distill line: DeepSeek-R1-Distill-Qwen-14B brings 14.8B dense parameters on a Qwen-2.5-14B base with full R1 reasoning structures into the Desktop range. 128,000 tokens of context, MIT license, locally as Unsloth-GGUF — the reasoning compromise between Edge suitability and Workstation capacity.
- Open Weights
- Desktop
- llama.cpp
- Text
- Instruction-Tuned
- Unusable
Sovereign Risk: MEDIUM TODO
Key metrics
Score · Latency · Cost · Quality
- Total Score Bronze
- 58.27
- Routine
- 37.16
- Reasoning
- 21.11
- LLM Judge Avg
- 2.55 / 5
- 100 Coverage
- Avg Task Duration
- 126.66s
- Unusable
- Token Rate
- 15.01tok/s
- Output Rate
- P95 Latency
- 159.69s
- Top 5 %
- Total Tokens
- 100500
- Output Volume
- Cost per 1K
- $0
- USD / 1K Requests
- Benchmark Cost
- $0
- Total · 100500 tok
Benchmark modules
10 modules · weighted · vs. model median & top performer
DeepSeek R1 Distill Qwen 14B
Best model
Ø All models
Code Quality
62.6
CLI Benchmark
71.67
Logical Reasoning
48.32
UX Writing
57.75
Documentation
57.63
Content Transform.
62.72
Cultural Intelligence
53.9
Synthesis Quality
Tool Execution
ToolUse Score
Benchmark Cost
$0
Token efficiency & latency
Consumption per module vs. model median