GLM-5.1

GLM-5.1 is Z.AI’s post-training upgrade with 754 billion total and 40 billion active parameters in a MoE architecture, optimized for long-horizon agentic coding workflows with up to eight hours of autonomous execution. The context window spans 200,000 tokens, and the weights are available as an Open Weights model under the MIT license.

Zhipu AI Version 5.1 Commercial use permitted MoE 754 B (40 B active) 200 K Context 12/2025 $1.05 / $3.5 per 1M

  • Open Weights
  • Frontier
  • OpenRouter
  • Text
  • Instruction-Tuned
  • Agentic Orchestrator
  • Batch

Sovereign Risk: HIGH Z.AI (formerly Zhipu AI) is a Chinese company and subject to China’s National Security Law (NSL), which can enable state access to data. In February 2025, Germany’s BSI explicitly warned against the use of Chinese AI cloud services (BSI reference: Warning DeepSeek, 04.02.2025); this risk assessment applies analogously to all Chinese cloud AI providers that process user data on Chinese servers. With purely local inference, the Cloud Act-equivalent risk does not apply.

Key metrics

Score · Latency · Cost · Quality

Total Score Silver
73.34
Routine
44.56
Reasoning
28.78

Rank #37

LLM Judge Avg
3.74
100 Coverage
Avg Task Duration
63.3
Batch
Token Rate
64.84
Output Rate
P95 Latency
192.19
Top 5 %
Total Tokens
149000
Output Volume
Cost per 1K
$0.0035
USD / 1K Requests
Benchmark Cost
$0.52
Total · 149000 tok

Benchmark modules

10 modules · weighted · vs. model median & top performer

GLM-5.1 Best model Ø All models
Code Quality 71.28
CLI Benchmark 93
Logical Reasoning 74.4
UX Writing 72.21
Documentation 69.88
Content Transform. 79.51
Cultural Intelligence 68.52
Synthesis Quality 59.17
Tool Execution 90
ToolUse Score 67.75
Benchmark Cost $0.52

Token efficiency & latency

Consumption per module vs. model median

Token consumption per module

Performance profile