Qwen3.8-2.4T-A95B

Qwen3.8-2.4T-A95B is Alibaba’s first Open Weights release at Qwen-Max level (August 12, 2026), a fine-grained Mixture-of-Experts model with 2.4 trillion total and 95 billion active parameters per token. License: proprietary ‘Qwen3.8-Max License’. The model processes text in a 262,144-token context (expandable to approximately one million) with a mandatory reasoning mode (low/high/xhigh) and hybrid attention combining Gated-DeltaNet and Gated-Attention.

Alibaba Version 3.8 Commercial use permitted MoE 2400 B (95 B active) 262 K Context $2 / $6 per 1M

  • Open Weights
  • Frontier
  • OpenRouter
  • Text
  • Long Context
  • Agentic Orchestrator
  • Unusable

Sovereign Risk: HIGH The model is developed by the Qwen Team at Alibaba Cloud, a company headquartered in China. Due to Chinese legislation (including the National Security Law) and the associated potential for state influence over technology companies, the origin risk of the weights is classified as high, regardless of the deployment location.

Key metrics

Score · Latency · Cost · Quality

Total Score Silver
76.48
Routine
46.82
Reasoning
29.67

Rank #23

LLM Judge Avg
3.79
100 Coverage
Avg Task Duration
121.65
Unusable
Token Rate
95.04
Output Rate
P95 Latency
445.42
Top 5 %
Total Tokens
370700
Output Volume
Cost per 1K
$0.006
USD / 1K Requests
Benchmark Cost
$2.22
Total · 370700 tok

Benchmark modules

10 modules · weighted · vs. model median & top performer

Qwen3.8-2.4T-A95B Best model Ø All models
Code Quality 82.04
CLI Benchmark 86.67
Logical Reasoning 69.8
UX Writing 72.79
Documentation 72.51
Content Transform. 77.35
Cultural Intelligence 74.64
Synthesis Quality 66.67
Tool Execution 90
ToolUse Score 81.17
Benchmark Cost $2.22

Token efficiency & latency

Consumption per module vs. model median

Token consumption per module

Performance profile