Qwen 3 14B

Qwen 3 14B is Alibaba’s open-weights model for general language tasks and reasoning with an optional thinking mode. The Q6 quantization is designed for efficient local operation without a cloud connection; the context window covers 128,000 tokens. Fully commercially usable under the Apache 2.0 license.

Alibaba Version 3 Commercial use permitted Dense 14 B (14 B active) 128 K Context 09/2024 locally tested

  • Open Weights
  • Desktop
  • llama.cpp
  • Text
  • Interactive

Sovereign Risk: LOW The model is operated locally without a cloud connection; CLOUD Act and data transfer risks do not apply to purely local inference. The sovereign risk refers to the weights provenance, not to active data transmission.

Key metrics

Score · Latency · Cost · Quality

Total Score Silver
67.66
Routine
41.11
Reasoning
26.55

Rank #79

LLM Judge Avg
3.4
100 Coverage
Avg Task Duration
34.3
Interactive
Token Rate
23.87
Output Rate
P95 Latency
87.96
Top 5 %
Total Tokens
55900
Output Volume
Cost per 1K
$0
USD / 1K Requests
Benchmark Cost
$0
Total · 55900 tok

Benchmark modules

10 modules · weighted · vs. model median & top performer

Qwen 3 14B Best model Ø All models
Code Quality 67.8
CLI Benchmark 91.67
Logical Reasoning 63.09
UX Writing 62.55
Documentation 60.85
Content Transform. 70.38
Cultural Intelligence 71.3
Synthesis Quality 41.67
Tool Execution 90
ToolUse Score 65.67
Benchmark Cost $0

Token efficiency & latency

Consumption per module vs. model median

Token consumption per module

Performance profile