GLM-5.3
GLM-5.3 is Z.AI’s Frontier coding flagship from August 14, 2026, with 744 billion total and 40 billion active parameters. All improvements over GLM-5.2 stem from post-training: Terminal-Bench 3.0 rose from 4.6 to 28.3 percent, ExploitBench more than doubled. Reasoning is mandatorily active, context one million tokens. Weights and license not yet released, Chinese cloud infrastructure.
- Open Weights
- Frontier
- OpenRouter
- Text
- Agentic Orchestrator
- Long Context
- Batch
Sovereign Risk: HIGH The developer Z.AI is headquartered in China. The model’s development is therefore subject to Chinese legislation, which represents an elevated risk in the international context with regard to data security and state influence. As of the current date (August 23, 2026), the weights have not yet been released; the model is exclusively available via the GLM Coding Plan and the Z.AI cloud infrastructure, meaning all requests are routed through Chinese servers. Risk mitigation through local deployment is not currently possible.
Key metrics
Score · Latency · Cost · Quality
- Total Score Silver
- 67.7
- Routine
- 38.09
- Reasoning
- 29.61
- LLM Judge Avg
- 3.49 / 5
- 100 Coverage
- Avg Task Duration
- 113.32s
- Batch
- Token Rate
- 53.66tok/s
- Output Rate
- P95 Latency
- 274.67s
- Top 5 %
- Total Tokens
- 334200
- Output Volume
- Cost per 1K
- $0.0044
- USD / 1K Requests
- Benchmark Cost
- $1.47
- Total · 334200 tok
Benchmark modules
10 modules · weighted · vs. model median & top performer
Token efficiency & latency
Consumption per module vs. model median