GLM-5.1
GLM-5.1 is Z.AI’s post-training upgrade with 754 billion total and 40 billion active parameters in a MoE architecture, optimized for long-horizon agentic coding workflows with up to eight hours of autonomous execution. The context window spans 200,000 tokens, and the weights are available as an Open Weights model under the MIT license.
- Open Weights
- Frontier
- OpenRouter
- Text
- Instruction-Tuned
- Agentic Orchestrator
- Batch
Sovereign Risk: HIGH Z.AI (formerly Zhipu AI) is a Chinese company and subject to China’s National Security Law (NSL), which can enable state access to data. In February 2025, Germany’s BSI explicitly warned against the use of Chinese AI cloud services (BSI reference: Warning DeepSeek, 04.02.2025); this risk assessment applies analogously to all Chinese cloud AI providers that process user data on Chinese servers. With purely local inference, the Cloud Act-equivalent risk does not apply.
Key metrics
Score · Latency · Cost · Quality
- Total Score Silver
- 73.34
- Routine
- 44.56
- Reasoning
- 28.78
- LLM Judge Avg
- 3.74 / 5
- 100 Coverage
- Avg Task Duration
- 63.3s
- Batch
- Token Rate
- 64.84tok/s
- Output Rate
- P95 Latency
- 192.19s
- Top 5 %
- Total Tokens
- 149000
- Output Volume
- Cost per 1K
- $0.0035
- USD / 1K Requests
- Benchmark Cost
- $0.52
- Total · 149000 tok
Benchmark modules
10 modules · weighted · vs. model median & top performer
Token efficiency & latency
Consumption per module vs. model median