GLM-4.7

GLM-4.7 is Zhipu AI’s flagship model with 355 billion total and 32 billion active parameters in a MoE architecture, optimized for agentic coding, reasoning, and bilingual tasks in Chinese and English. The model supports a switchable thinking system and is available as an Open Weights variant for local deployment or via cloud interfaces.

Zhipu AI Version 4.7 Commercial use permitted MoE 355 B (32 B active) 128 K Context 12/2025 $0.38 / $1.74 per 1M

  • Restricted Weights
  • Frontier
  • OR
  • Text
  • Instruction-Tuned
  • Batch

Sovereign Risk: HIGH Zhipu AI / Z.AI is a Chinese company and subject to China’s National Security Law (NSL), which may allow state access to data. In February 2025, Germany’s BSI explicitly warned against the use of Chinese AI cloud services (BSI reference: Warning DeepSeek, 04.02.2025); this risk assessment applies analogously to all Chinese cloud AI providers that process user data on Chinese servers. With purely local inference, the Cloud Act-equivalent risk does not apply.

Key metrics

Score · Latency · Cost · Quality

Total Score Silver
74.47
Routine
45.16
Reasoning
29.32

Rank #27

LLM Judge Avg
3.9
100 Coverage
Avg Task Duration
48.84
Batch
Token Rate
26.12
Output Rate
P95 Latency
122.18
Top 5 %
Total Tokens
145700
Output Volume
Cost per 1K
$0.0017
USD / 1K Requests
Benchmark Cost
$0.25
Total · 145700 tok

Benchmark modules

10 modules · weighted · vs. model median & top performer

GLM-4.7 Best model Ø All models
Code Quality 70.56
CLI Benchmark 91.34
Logical Reasoning 76.87
UX Writing 75.73
Documentation 67.87
Content Transform. 77.14
Cultural Intelligence 81.72
Synthesis Quality 42.5
Tool Execution 83.33
ToolUse Score 62.75
Benchmark Cost $0.25

Token efficiency & latency

Consumption per module vs. model median

Token consumption per module

Performance profile