GLM 4.6

GLM-4.6 is Zhipu AI’s Frontier language model with a focus on Chinese and English language proficiency. The model operates with a context window of 128,000 tokens and supports tool use for agentic workflows. Due to the Chinese manufacturer jurisdiction, a separate privacy assessment is required, and commercial use is subject to restrictions.

Zhipu AI Version 4.6 Commercial use restricted MoE 355 B (32 B active) 128 K Context 06/2025 $0.43 / $1.75 per 1M

  • Restricted Weights
  • Server
  • OpenRouter
  • Text
  • Instruction-Tuned
  • Batch

Sovereign Risk: HIGH Zhipu AI is a Chinese company and subject to China’s National Security Law (NSL), which may allow state access to data. In February 2025, the BSI explicitly warned against the use of Chinese AI cloud services (BSI reference: Warning DeepSeek, 04.02.2025); this risk assessment applies analogously to all Chinese cloud AI providers that process user data on Chinese servers.

Key metrics

Score · Latency · Cost · Quality

Total Score Silver
76.17
Routine
47.12
Reasoning
29.05

Rank #16

LLM Judge Avg
3.85
100 Coverage
Avg Task Duration
69.56
Batch
Token Rate
46.1
Output Rate
P95 Latency
171.17
Top 5 %
Total Tokens
145000
Output Volume
Cost per 1K
$0.0018
USD / 1K Requests
Benchmark Cost
$0.25
Total · 145000 tok

Benchmark modules

10 modules · weighted · vs. model median & top performer

GLM 4.6 Best model Ø All models
Code Quality 72.6
CLI Benchmark 89
Logical Reasoning 71.63
UX Writing 71.35
Documentation 74.43
Content Transform. 80.23
Cultural Intelligence 81.32
Synthesis Quality 63.33
Tool Execution 89.17
ToolUse Score 53.17
Benchmark Cost $0.25

Token efficiency & latency

Consumption per module vs. model median

Token consumption per module

Performance profile