Qwen 2.5 Coder 7B
Qwen 2.5 Coder 7B is a compact Open Weights coding model by Alibaba, optimized for code generation, debugging, and repair. Q6 quantization enables local operation on resource-efficient hardware with minimal quality loss; the context window spans 128,000 tokens for complex codebases. Fully commercially usable under the Apache 2.0 license.
- Open Weights
- Nano
- llama.cpp
- Text
- Instruction-Tuned
- Real-Time
Sovereign Risk: LOW Fully local inference without cloud connectivity. The weights are publicly available (Apache 2.0) and run entirely locally. NSL is not relevant, as no data is transmitted to Alibaba infrastructure.
Key metrics
Score · Latency · Cost · Quality
- Total Score Bronze
- 56.09
- Routine
- 33.79
- Reasoning
- 22.31
- LLM Judge Avg
- 2.74 / 5
- 100 Coverage
- Avg Task Duration
- 16.93s
- Real-Time
- Token Rate
- 50.07tok/s
- Output Rate
- P95 Latency
- 31.85s
- Top 5 %
- Total Tokens
- 54100
- Output Volume
- Cost per 1K
- $0
- USD / 1K Requests
- Benchmark Cost
- $0
- Total · 54100 tok
Benchmark modules
10 modules · weighted · vs. model median & top performer
Token efficiency & latency
Consumption per module vs. model median