Qwen 2.5 Coder 7B

A Q6_K-GGUF distribution of Qwen 2.5 Coder 7B for local coding: 7.6 billion dense parameters, Apache-2.0 license, specialized in code generation, debugging, and repair. The family supports 128,000 tokens of context; in GGUF setups, only 32,000 are natively available without long-context configuration. Fully commercially usable, compact and efficient on Workstation hardware.

Alibaba Version 2.5 Commercial use permitted Dense 7.61 B 128 K Context 09/2024 locally tested

  • Open Weights
  • Edge
  • llama.cpp
  • Text
  • Instruction-Tuned
  • Interactive

Sovereign Risk: LOW The weights originate from Alibaba’s Apache-2.0-licensed Qwen2.5-Coder family and are run entirely locally here. Without a cloud connection, operational risk is low; the provenance remains Chinese-jurisdictional, but local Open Weights usage minimizes data exposure.[web:875][web:876][web:878]

Key metrics

Score · Latency · Cost · Quality

Total Score Bronze
55.32
Routine
33.87
Reasoning
21.45

Rank #104

LLM Judge Avg
2.53
100 Coverage
Avg Task Duration
33.7
Interactive
Token Rate
37.43
Output Rate
P95 Latency
46.54
Top 5 %
Total Tokens
68800
Output Volume
Cost per 1K
$0
USD / 1K Requests
Benchmark Cost
$0
Total · 68800 tok

Benchmark modules

10 modules · weighted · vs. model median & top performer

Qwen 2.5 Coder 7B Best model Ø All models
Code Quality 52.3
CLI Benchmark 85.56
Logical Reasoning 53.39
UX Writing 54.75
Documentation 50.2
Content Transform. 58.92
Cultural Intelligence 44.6
Synthesis Quality 34.17
Tool Execution 82.5
ToolUse Score 57.92
Benchmark Cost $0

Token efficiency & latency

Consumption per module vs. model median

Token consumption per module

Performance profile