Qwen3.8-2.4T-A95B
Qwen3.8-2.4T-A95B is Alibaba’s first Open Weights release at Qwen-Max level (August 12, 2026), a fine-grained Mixture-of-Experts model with 2.4 trillion total and 95 billion active parameters per token. License: proprietary ‘Qwen3.8-Max License’. The model processes text in a 262,144-token context (expandable to approximately one million) with a mandatory reasoning mode (low/high/xhigh) and hybrid attention combining Gated-DeltaNet and Gated-Attention.
- Open Weights
- Frontier
- OpenRouter
- Text
- Long Context
- Agentic Orchestrator
- Batch
Sovereign Risk: HIGH The model is developed by the Qwen Team at Alibaba Cloud, a company headquartered in China. Due to Chinese legislation (including the National Security Law) and the associated potential for state influence over technology companies, the origin risk of the weights is classified as high, regardless of the deployment location.
Key metrics
Score · Latency · Cost · Quality
- Total Score Silver
- 67.23
- Routine
- 39.91
- Reasoning
- 27.31
- LLM Judge Avg
- 3.43 / 5
- 100 Coverage
- Avg Task Duration
- 73.94s
- Batch
- Token Rate
- 45.14tok/s
- Output Rate
- P95 Latency
- 239.68s
- Top 5 %
- Total Tokens
- 315700
- Output Volume
- Cost per 1K
- $0.006
- USD / 1K Requests
- Benchmark Cost
- $1.89
- Total · 315700 tok
Benchmark modules
10 modules · weighted · vs. model median & top performer
Token efficiency & latency
Consumption per module vs. model median