GLM 4.6

GLM-4.6 is Zhipu AI’s Frontier language model with a focus on Chinese and English language proficiency. The model operates with a context window of 128,000 tokens and supports tool use for agentic workflows. Due to the Chinese manufacturer jurisdiction, a separate privacy assessment is required, and commercial use is subject to restrictions.

Zhipu AI Version 4.6 Commercial use restricted Dense 128 K Context 06/2025 $0.39 / $1.9 per 1M

  • Restricted Weights
  • Frontier
  • OR
  • Text
  • Instruction-Tuned
  • Batch

Sovereign Risk: HIGH Zhipu AI is a Chinese company and subject to China’s National Security Law (NSL), which may allow state access to data. In February 2025, the BSI explicitly warned against the use of Chinese AI cloud services (BSI reference: Warning DeepSeek, 04.02.2025); this risk assessment applies analogously to all Chinese cloud AI providers that process user data on Chinese servers.

Key metrics

Score · Latency · Cost · Quality

Total Score Silver
73.14
Routine
43.8
Reasoning
29.33

Rank #45

LLM Judge Avg
3.67
100 Coverage
Avg Task Duration
63.18
Batch
Token Rate
17.8
Output Rate
P95 Latency
152.26
Top 5 %
Total Tokens
139900
Output Volume
Cost per 1K
$0.0019
USD / 1K Requests
Benchmark Cost
$0.27
Total · 139900 tok

Benchmark modules

10 modules · weighted · vs. model median & top performer

GLM 4.6 Best model Ø All models
Code Quality 69.32
CLI Benchmark 78.67
Logical Reasoning 76.46
UX Writing 65.07
Documentation 72.08
Content Transform. 72.95
Cultural Intelligence 78.52
Synthesis Quality 63.33
Tool Execution 89.17
ToolUse Score 75.92
Benchmark Cost $0.27

Token efficiency & latency

Consumption per module vs. model median

Token consumption per module

Performance profile