DeepSeek R1 Distill Qwen 14B

The strongest variant in the DeepSeek-R1-Distill line: DeepSeek-R1-Distill-Qwen-14B brings 14.8B dense parameters on a Qwen-2.5-14B base with full R1 reasoning structures into the Desktop range. 128,000 tokens of context, MIT license, locally as Unsloth-GGUF — the reasoning compromise between Edge suitability and Workstation capacity.

DeepSeek Version 1 Commercial use permitted Dense 14.8 B (14.8 B active) 128 K Context 06/2024 locally tested

  • Open Weights
  • Desktop
  • llama.cpp
  • Text
  • Instruction-Tuned
  • Unusable

Sovereign Risk: MEDIUM TODO

Key metrics

Score · Latency · Cost · Quality

Total Score Bronze
58.27
Routine
37.16
Reasoning
21.11

Rank #96

LLM Judge Avg
2.55
100 Coverage
Avg Task Duration
126.66
Unusable
Token Rate
15.01
Output Rate
P95 Latency
159.69
Top 5 %
Total Tokens
100500
Output Volume
Cost per 1K
$0
USD / 1K Requests
Benchmark Cost
$0
Total · 100500 tok

Benchmark modules

10 modules · weighted · vs. model median & top performer

DeepSeek R1 Distill Qwen 14B Best model Ø All models
Code Quality 62.6
CLI Benchmark 71.67
Logical Reasoning 48.32
UX Writing 57.75
Documentation 57.63
Content Transform. 62.72
Cultural Intelligence 53.9
Synthesis Quality
Tool Execution
ToolUse Score
Benchmark Cost $0

Token efficiency & latency

Consumption per module vs. model median

Token consumption per module

Performance profile