GLM 4.6
GLM-4.6 is Zhipu AI’s Frontier language model with a focus on Chinese and English language proficiency. The model operates with a context window of 128,000 tokens and supports tool use for agentic workflows. Due to the Chinese manufacturer jurisdiction, a separate privacy assessment is required, and commercial use is subject to restrictions.
- Restricted Weights
- Frontier
- OR
- Text
- Instruction-Tuned
- Batch
Sovereign Risk: HIGH Zhipu AI is a Chinese company and subject to China’s National Security Law (NSL), which may allow state access to data. In February 2025, the BSI explicitly warned against the use of Chinese AI cloud services (BSI reference: Warning DeepSeek, 04.02.2025); this risk assessment applies analogously to all Chinese cloud AI providers that process user data on Chinese servers.
Key metrics
Score · Latency · Cost · Quality
- Total Score Silver
- 73.14
- Routine
- 43.8
- Reasoning
- 29.33
- LLM Judge Avg
- 3.67 / 5
- 100 Coverage
- Avg Task Duration
- 63.18s
- Batch
- Token Rate
- 17.8tok/s
- Output Rate
- P95 Latency
- 152.26s
- Top 5 %
- Total Tokens
- 139900
- Output Volume
- Cost per 1K
- $0.0019
- USD / 1K Requests
- Benchmark Cost
- $0.27
- Total · 139900 tok
Benchmark modules
10 modules · weighted · vs. model median & top performer
Token efficiency & latency
Consumption per module vs. model median