Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Deep Infra logo

Model details

GLM-4.7

GLM-4.7 was unveiled under the title "GLM-4.7: Advancing the Coding Capability," a Z.ai release that frames the model around coding agents, complex reasoning, and tool use rather than as a general chat assistant. Community discussion around the announcement describes it as a mixture-of-experts design with 358 billion total parameters and 32 billion active per forward pass, a configuration that signals heavy routing capacity tuned for multi-step agent workflows rather than lightweight on-device inference. The same commentary highlights English and Chinese bilingual coverage and an OpenAI-style function-calling interface, suggesting the model is positioned for drop-in use inside existing agent frameworks and IDE integrations that already speak that schema.

Practically, GLM-4.7 appears aimed at developers building coding agents, retrieval-augmented pipelines, and tool-using assistants that need strong reasoning at long horizons. Its MoE sizing implies it will primarily run on well-provisioned servers rather than edge hardware, while the OpenAI-style tool calling convention should make it straightforward to swap into existing agent stacks. Distribution was still uneven shortly after launch: a developer forum thread from January 2026 noted that the model had not appeared on one major vendor catalog past a previously communicated mid-January target, hinting that ecosystem support across inference providers was catching up to the release rather than arriving on day one.

Deep Infrazai-org/GLM-4.7glm

Quick Info

Powered by
Provider
Deep Infra
Model key
zai-org/GLM-4.7
Release date
Dec 22, 2025
Last updated
Dec 22, 2025
Knowledge cutoff
2025-04
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.40
Output token cost
$1.75

Limits

Output tokens
16,384 tokens
Context window
202,752 tokens

Transparent token rates

Compare GLM-4.7 pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about GLM-4.7

No articles yet. Fetch the latest news to show it here.

Videos about GLM-4.7

More models around GLM-4.7