Qwen 2.5 Coder 7B
A Q6_K-GGUF distribution of Qwen 2.5 Coder 7B for local coding: 7.6 billion dense parameters, Apache-2.0 license, specialized in code generation, debugging, and repair. The family supports 128,000 tokens of context; in GGUF setups, only 32,000 are natively available without long-context configuration. Fully commercially usable, compact and efficient on Workstation hardware.
- Open Weights
- Edge
- llama.cpp
- Text
- Instruction-Tuned
- Interactive
Sovereign Risk: LOW The weights originate from Alibaba’s Apache-2.0-licensed Qwen2.5-Coder family and are run entirely locally here. Without a cloud connection, operational risk is low; the provenance remains Chinese-jurisdictional, but local Open Weights usage minimizes data exposure.[web:875][web:876][web:878]
Key metrics
Score · Latency · Cost · Quality
- Total Score Bronze
- 55.32
- Routine
- 33.87
- Reasoning
- 21.45
- LLM Judge Avg
- 2.53 / 5
- 100 Coverage
- Avg Task Duration
- 33.7s
- Interactive
- Token Rate
- 37.43tok/s
- Output Rate
- P95 Latency
- 46.54s
- Top 5 %
- Total Tokens
- 68800
- Output Volume
- Cost per 1K
- $0
- USD / 1K Requests
- Benchmark Cost
- $0
- Total · 68800 tok
Benchmark modules
10 modules · weighted · vs. model median & top performer
Token efficiency & latency
Consumption per module vs. model median