Qwen 3 Coder Next
Qwen 3 Coder Next is a coding-specialized Open Weights MoE model by Alibaba with 80 billion total and 3 billion active parameters. Q4 quantization significantly reduces memory requirements for local inference; the context window spans 262,000 tokens. Deployable locally on Workstation hardware under the Apache 2.0 license, optimized for coding agents and large codebases.
- Open Weights
- Workstation
- llama.cpp
- Text
- Instruction-Tuned
- Agentic Orchestrator
- Interactive
Key metrics
Score · Latency · Cost · Quality
- Total Score Silver
- 74.07
- Routine
- 45.22
- Reasoning
- 28.85
- LLM Judge Avg
- 3.67 / 5
- 100 Coverage
- Avg Task Duration
- 23.86s
- Interactive
- Token Rate
- 48.49tok/s
- Output Rate
- P95 Latency
- 65.8s
- Top 5 %
- Total Tokens
- 75100
- Output Volume
- Cost per 1K
- $0
- USD / 1K Requests
- Benchmark Cost
- $0
- Total · 75100 tok
Benchmark modules
10 modules · weighted · vs. model median & top performer
Token efficiency & latency
Consumption per module vs. model median