Sulat.com
AI models
Privatemode AI logo

Model details

GLM-5.3

GLM-5.3 is an open-weights model whose improvements over GLM-5.2 come entirely from scaled post-training rather than changes to the underlying base model. The training stack carried over from GLM-5.2 combines long-context infrastructure and asynchronous large-scale reinforcement learning on long-horizon task environments that the team has been building up over time. The release notes describe a deliberate push to spend more compute across more diverse environments and tasks, with the goal of sharpening the model on complex programming and multi-step work.

In practice, the model is positioned as a coding-focused system with notable emergent cyber capabilities. It shows a 50% gain over GLM-5.2 on the in-house Z.ai Code Bench and reaches open-source state-of-the-art results on Terminal Bench 3.0 and Agents' Last Exam, suggesting strong fit for real-world coding agents and long-running tool use. As post-training was scaled, vulnerability-discovery performance became state of the art on CyberGym, with the largest gains appearing further along the exploitation chain, indicating a model that is most useful where tasks demand extended, multi-stage reasoning rather than short answers.

Privatemode AIglm-5.3glmbeta

Quick Info

Powered by
Provider
Privatemode AI
Model key
glm-5.3
Release date
Aug 14, 2026
Last updated
Aug 14, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$1.791
Output token cost
$8.9436

Limits

Output tokens
131,072 tokens
Context window
256,000 tokens

Transparent token rates

Compare GLM-5.3 pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about GLM-5.3

Privatemode AI

CoverageBenchmark

Z.ai released GLM-5.3 on 14 August 2026, according to a launch-partner technical brief published the same day by Qubrid AI. GLM-5.3 uses the same 743B-parameter base as GLM-5.2, with all reported capability gains coming from scaled post-training (more executable environments, more environment types, and longer RL runs) GLM-5.3 ships with three reasoning effort levels (low, high, max, with max as default) and disabling thinking is no longer supported. Reported benchmark jumps include Terminal-Bench 3.0 rising from 4.6 to 28.3, DeepSWE v1.1 from 46.2 to 66.9, CyberGym at 84.5% (top of Z.ai's launch chart), and GDPval-AA v2 climbing fro

Privatemode AI

CoverageBenchmark

InferenceX's technical profile of GLM-5.3 quotes Z.ai's launch blog and docs directly, confirming that GLM-5.3 was announced on 14 August 2026 and is a post-training-only release on the same base as GLM-5.2, where "every gain comes from post-training." The model is framed as "Frontier Coding with Emergent Cyber Capabil GLM-5.3 is text-only and always reasoning, with a 1M-token context window and a 128K max output, and exposes three thinking-effort levels (low, high, max, default max), with the option to disable thinking no longer supported. The page also covers GLM-5.2 as the shared-base predecessor, noting the 1M-context long-horizo

Videos about GLM-5.3

More models around GLM-5.3