Qwen 3 Coder Next

Qwen 3 Coder Next is a coding-specialized Open Weights MoE model by Alibaba with 80 billion total and 3 billion active parameters. Q4 quantization significantly reduces memory requirements for local inference; the context window spans 262,000 tokens. Deployable locally on Workstation hardware under the Apache 2.0 license, optimized for coding agents and large codebases.

Alibaba Version 3 Coder Commercial use permitted MoE 80 B (3 B active) 262 K Context 05/2025 locally tested

  • Open Weights
  • Workstation
  • llama.cpp
  • Text
  • Instruction-Tuned
  • Agentic Orchestrator
  • Interactive

Key metrics

Score · Latency · Cost · Quality

Total Score Silver
74.07
Routine
45.22
Reasoning
28.85

Rank #33

LLM Judge Avg
3.67
100 Coverage
Avg Task Duration
23.86
Interactive
Token Rate
48.49
Output Rate
P95 Latency
65.8
Top 5 %
Total Tokens
75100
Output Volume
Cost per 1K
$0
USD / 1K Requests
Benchmark Cost
$0
Total · 75100 tok

Benchmark modules

10 modules · weighted · vs. model median & top performer

Qwen 3 Coder Next Best model Ø All models
Code Quality 79.8
CLI Benchmark 88.89
Logical Reasoning 64.64
UX Writing 71.65
Documentation 68.1
Content Transform. 75.73
Cultural Intelligence 80.3
Synthesis Quality 51.67
Tool Execution 90
ToolUse Score 70.88
Benchmark Cost $0

Token efficiency & latency

Consumption per module vs. model median

Token consumption per module

Performance profile