Qwen 3 14B
Qwen 3 14B is Alibaba’s open-weights model for general language tasks and reasoning with an optional thinking mode. The Q6 quantization is designed for efficient local operation without a cloud connection; the context window covers 128,000 tokens. Fully commercially usable under the Apache 2.0 license.
- Open Weights
- Desktop
- llama.cpp
- Text
- Interactive
Sovereign Risk: LOW The model is operated locally without a cloud connection; CLOUD Act and data transfer risks do not apply to purely local inference. The sovereign risk refers to the weights provenance, not to active data transmission.
Key metrics
Score · Latency · Cost · Quality
- Total Score Silver
- 67.66
- Routine
- 41.11
- Reasoning
- 26.55
- LLM Judge Avg
- 3.4 / 5
- 100 Coverage
- Avg Task Duration
- 34.3s
- Interactive
- Token Rate
- 23.87tok/s
- Output Rate
- P95 Latency
- 87.96s
- Top 5 %
- Total Tokens
- 55900
- Output Volume
- Cost per 1K
- $0
- USD / 1K Requests
- Benchmark Cost
- $0
- Total · 55900 tok
Benchmark modules
10 modules · weighted · vs. model median & top performer
Token efficiency & latency
Consumption per module vs. model median