Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Inceptron logo

Model details

GLM 5.2

GLM-5.2 is the base language model released by z.ai that anchors a larger family of post-trained variants, most notably GLM-5.3, which explicitly reuses the GLM-5.2 base and attributes all of its gains to scaled post-training rather than a new pretraining run. The release is framed as the point at which z.ai assembled a long-context and long-horizon training stack, combining IndexShare for efficient long-context processing, SAO for reinforcement learning on long-horizon tasks, and the slime framework for large-scale asynchronous training. An NVIDIA NGC catalog listing under the zai-org team namespace also places the model within the NVIDIA NIM ecosystem, suggesting enterprise-grade serving integration alongside its public availability.

In practical terms, GLM-5.2 functions as a foundation for advanced coding and agentic workloads rather than as a turnkey product. z.ai uses it as the substrate for GLM-5.3, which reportedly delivers a fifty percent improvement on the in-house Z.ai Code Bench over GLM-5.2 and reaches open-source state-of-the-art on Terminal Bench 3.0 and Agents' Last Exam, with emergent strengths on CyberGym that compound further up the exploitation chain. For teams choosing GLM-5.2 directly, the model is best understood as a versatile open base suitable for further fine-tuning, long-context applications, and tool-augmented pipelines, while those who want the latest coding and cyber capability should evaluate the successor variant built on top of it.

Inceptronzai-org/GLM-5.2glm

Quick Info

Powered by
Provider
Inceptron
Model key
zai-org/GLM-5.2
Release date
Jun 13, 2026
Last updated
Jun 13, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.71
Output token cost
$2.35

Limits

Output tokens
1,048,576 tokens
Context window
1,048,576 tokens

Transparent token rates

Compare GLM 5.2 pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about GLM 5.2

CoreWeave

CoverageBenchmark

On June 17, 2026, Z.ai published the GLM 5.2 benchmark scorecard that was withheld at the June 13 launch, alongside MIT-licensed open weights for both zai-org/GLM-5.2 and zai-org/GLM-5.2-FP8 on HuggingFace — arriving earlier than the originally promised "the following week" timeline. GLM 5.2 posts 62.1 on SWE-bench Pro The article frames GLM 5.2 as the first credibly open-weight model to lead an Anthropic or OpenAI flagship on a real-world SWE-bench Pro head-to-head, noting that teams using Claude Code, the Claude Agent SDK, Cursor, or the Vercel AI Gateway now have a frontier-tier open-weight backend they can self-host. It also reca

Lilac

CoverageRelease Notes

ThursdAI's June 2026 monthly roundup lists GLM-5.2 as one of 32 AI releases covered that month, attributing it to Z.ai (Zhipu AI) and categorizing it for developers and coding agents. The aggregator reports concrete model-intrinsic specs: a 753-billion-parameter open Mixture-of-Experts architecture and a 1M-token conte As a secondary podcast/newsletter aggregator, ThursdAI corroborates Z.ai authorship and the 753B open-MoE / 1M-context framing, but its single-source GPQA Diamond figure and other benchmark numbers are not independently verified against a Z.ai release post or model card within this candidate set. The roundup neverthele

CoreWeave

CoverageBenchmark

Z.ai (formerly Zhipu AI / THUDM) released GLM-5.2 on June 13, 2026, as a 744-billion-parameter open-weight Mixture-of-Experts model under the MIT license, activating roughly 40B parameters per token across 384 experts and supporting a 1,000,000-token context window with a 131,072-token maximum output. The article docum Technically, GLM-5.2 introduces IndexShare sparse attention, which the source cites as achieving a 2.9x FLOP reduction at the full 1M context length, along with an improved Multi-Token Prediction speculative decoding path. The model uses a higher activation ratio than Kimi K2.7, and the article walks through local depl

Weights & Biases

Coverage

The U.S. National Institute of Standards and Technology's Center for AI Standards and Innovation (CAISI) published an independent assessment of Z.ai's open-weight GLM-5.2 model on July 8, 2026, roughly three weeks after its June 16, 2026 release. CAISI concluded that GLM-5.2 was probably the most capable open-weight AI On safeguards and security, CAISI found mixed results: GLM-5.2's safeguards permitted assistance with agentic cyber exploit development and blocked fewer sensitive biological questions than reference U.S. models, but it appeared potentially more robust than other evaluated PRC open-weight models against agent-hijacking

Lilac

CoverageAnalysis

A Hacker News discussion (916 points, ~80 days old) centered on Artificial Analysis's leaderboard post declaring GLM-5.2 the new leading open-weights model. Commenters provide concrete behavioral observations: GLM 5.2 supports reasoning-effort tiers ("high" and "xhigh," with xhigh mapped to max effort), and at xhigh it The thread frames GLM 5.2 as a significant step up for open weights and getting close to frontier, while flagging reasoning efficiency as the next bottleneck relative to GPT 5.5. It also reinforces Z.ai's open-weights positioning by emphasizing the cost gap (GLM 5.2 expected to undercut Opus 4.8 and GPT 5.5 on price ev

Videos about GLM 5.2

More models around GLM 5.2