GLM-5.3

GLM-5.3 is Z.AI’s Frontier coding flagship from August 14, 2026, with 744 billion total and 40 billion active parameters. All improvements over GLM-5.2 stem from post-training: Terminal-Bench 3.0 rose from 4.6 to 28.3 percent, ExploitBench more than doubled. Reasoning is mandatorily active, context one million tokens. Weights and license not yet released, Chinese cloud infrastructure.

Zhipu AI Version 5.3 Commercial use restricted MoE 744 B (40 B active) 1000 K Context 04/2026 $1.4 / $4.4 per 1M

  • Open Weights
  • Frontier
  • OpenRouter
  • Text
  • Agentic Orchestrator
  • Long Context
  • Batch

Sovereign Risk: HIGH The developer Z.AI is headquartered in China. The model’s development is therefore subject to Chinese legislation, which represents an elevated risk in the international context with regard to data security and state influence. As of the current date (August 23, 2026), the weights have not yet been released; the model is exclusively available via the GLM Coding Plan and the Z.AI cloud infrastructure, meaning all requests are routed through Chinese servers. Risk mitigation through local deployment is not currently possible.

Key metrics

Score · Latency · Cost · Quality

Total Score Silver
67.7
Routine
38.09
Reasoning
29.61

Rank #78

LLM Judge Avg
3.49
100 Coverage
Avg Task Duration
113.32
Batch
Token Rate
53.66
Output Rate
P95 Latency
274.67
Top 5 %
Total Tokens
334200
Output Volume
Cost per 1K
$0.0044
USD / 1K Requests
Benchmark Cost
$1.47
Total · 334200 tok

Benchmark modules

10 modules · weighted · vs. model median & top performer

GLM-5.3 Best model Ø All models
Code Quality 77.24
CLI Benchmark 78.67
Logical Reasoning 76.33
UX Writing 44.63
Documentation 54.2
Content Transform. 58.02
Cultural Intelligence 78.52
Synthesis Quality 73.33
Tool Execution 90
ToolUse Score 79.5
Benchmark Cost $1.47

Token efficiency & latency

Consumption per module vs. model median

Token consumption per module

Performance profile