Qwen3.8-2.4T-A95B

Qwen3.8-2.4T-A95B is Alibaba’s first Open Weights release at Qwen-Max level (August 12, 2026), a fine-grained Mixture-of-Experts model with 2.4 trillion total and 95 billion active parameters per token. License: proprietary ‘Qwen3.8-Max License’. The model processes text in a 262,144-token context (expandable to approximately one million) with a mandatory reasoning mode (low/high/xhigh) and hybrid attention combining Gated-DeltaNet and Gated-Attention.

Qwen Version 3.8 Commercial use permitted MoE 2400 B (95 B active) 262 K Context $2 / $6 per 1M

  • Open Weights
  • Frontier
  • OpenRouter
  • Text
  • Long Context
  • Agentic Orchestrator
  • Batch

Sovereign Risk: HIGH The model is developed by the Qwen Team at Alibaba Cloud, a company headquartered in China. Due to Chinese legislation (including the National Security Law) and the associated potential for state influence over technology companies, the origin risk of the weights is classified as high, regardless of the deployment location.

Key metrics

Score · Latency · Cost · Quality

Total Score Silver
67.23
Routine
39.91
Reasoning
27.31

Rank #78

LLM Judge Avg
3.43
100 Coverage
Avg Task Duration
73.94
Batch
Token Rate
45.14
Output Rate
P95 Latency
239.68
Top 5 %
Total Tokens
315700
Output Volume
Cost per 1K
$0.006
USD / 1K Requests
Benchmark Cost
$1.89
Total · 315700 tok

Benchmark modules

10 modules · weighted · vs. model median & top performer

Qwen3.8-2.4T-A95B Best model Ø All models
Code Quality 61.8
CLI Benchmark 84.34
Logical Reasoning 67.17
UX Writing 53.27
Documentation 46.12
Content Transform. 77.83
Cultural Intelligence 77.84
Synthesis Quality 66.67
Tool Execution 90
ToolUse Score 78
Benchmark Cost $1.89

Token efficiency & latency

Consumption per module vs. model median

Token consumption per module

Performance profile