GLM 5.3
Specifications
- Input
- Output
- Context window
- 1M tokens
- Released
- Aug 2026
Performance
- Speed
- 174 t/s
- TTFT
- 856 ms
- Latency
- 451 ms
- Intelligence
- —
Pricing
- Input
- €1.00 per 1M tokens
- Output
- €4.00 per 1M tokens
€0.25 cache hit
About this model
ZAI GLM 5.3 is a 744B parameter Mixture-of-Experts (MoE) language model built on the GLM-MoE DSA architecture, using the same base model as GLM-5.2 with significant post-training improvements focused on complex coding and long-horizon agentic tasks. It features a 1-million token context window and configurable reasoning effort levels (low, high, max), with thinking enabled by default. The model achieves strong benchmark results including 88.2 on Terminal Bench 2.1, 66.9 on DeepSWE (v1.1), 62.5 on HLE with Tools, and 73.0 on Toolathlon Verified. It is released under a custom license.
Technical specifications
- Capabilities
- Input modalities
- Output modalities
- Reasoning
- Hybrid Default on
Knowledge horizon
Released Aug 2026
Today
Since release 0 mo
See also
Add Model to Comparison
Search for a model to add