Command Palette
Search for a command to run

GLM 5.3

by ZAI

Specifications

Input
Output
Context window
1M tokens
Released
Aug 2026

Performance

Speed
174 t/s
TTFT
856 ms
Latency
451 ms
Intelligence

Pricing

Input
€1.00
per 1M tokens
Output
€4.00
per 1M tokens
€0.25 cache hit

About this model

ZAI GLM 5.3 is a 744B parameter Mixture-of-Experts (MoE) language model built on the GLM-MoE DSA architecture, using the same base model as GLM-5.2 with significant post-training improvements focused on complex coding and long-horizon agentic tasks. It features a 1-million token context window and configurable reasoning effort levels (low, high, max), with thinking enabled by default. The model achieves strong benchmark results including 88.2 on Terminal Bench 2.1, 66.9 on DeepSWE (v1.1), 62.5 on HLE with Tools, and 73.0 on Toolathlon Verified. It is released under a custom license.

Technical specifications

Capabilities
Input modalities
Output modalities
Reasoning
Hybrid Default on

Knowledge horizon

Released Aug 2026
Today
Since release 0 mo

See also

Add Model to Comparison
Search for a model to add
Command Palette
Search for a command to run