Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
NovitaAI logo

Model details

GLM-5

GLM-5 is a large open-weights mixture-of-experts model positioned for complex systems engineering and long-horizon agentic work rather than casual code generation. The architecture scales from the prior generation's 355B parameters with 32B active to 744B parameters with 40B active, supported by a pre-training corpus expanded from 23T to 28.5T tokens. To keep deployment economical at this scale, GLM-5 integrates DeepSeek Sparse Attention, which compresses long-context computation without sacrificing the model's ability to handle extended inputs. Together, these changes mark a deliberate step from snippet-level code assistance toward end-to-end engineering, where the model is expected to plan, coordinate tools, and sustain multi-step tasks across long contexts.

On the post-training side, the team paired a new asynchronous reinforcement-learning infrastructure called slime with continued scale, enabling more granular iterations on long-horizon task environments. The result is broad gains on academic benchmarks relative to earlier GLM variants, with the model also surfaced through developer-friendly channels such as Novita AI's Model APIs and Amazon Bedrock, confirming production availability for teams that want to wire it into agentic pipelines. Practical fit is strongest for organizations building software agents, automation backends, or research prototypes that need a capable, open-weights base they can self-host or route through managed inference, and who value long-context reasoning paired with sparse-attention efficiency over a smaller, denser alternative.

NovitaAIzai-org/glm-5glm

Quick Info

Powered by
Provider
NovitaAI
Model key
zai-org/glm-5
Release date
Feb 11, 2026
Last updated
Feb 12, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$1.00
Output token cost
$3.20

Limits

Output tokens
131,072 tokens
Context window
202,800 tokens

Transparent token rates

Compare GLM-5 pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about GLM-5

NovitaAI

CoverageRelease Notes

Discover more about what's new at AWS with Minimax M2.5 and GLM 5 models now available on Amazon Bedrock

NovitaAI

Coverage

SINGAPORE, February 16, 2026--GLM-5, newly released as open source, signals a broader shift in artificial intelligence. Large language models are moving beyond generating code snippets or interface prototypes toward building complete systems and carrying out complex, end-to-end tasks. The change marks a transition from

NovitaAI

Official sourceAnalysis

Build with GLM-5 on Novita AI. 744B MoE architecture for expert agentic coding. Claim your free credits and start building today!, Post a Comment

NovitaAI

CoverageRelease Notes

Big-AGI's public changelog lists "Z.ai GLM-5" support added alongside Moonshot Kimi K2.7-code on June 16, 2026, confirming that the Z.ai GLM-5 model is generally reachable through the OpenRouter-style routing that consumer apps aggregate. The changelog entry indicates the model is exposed as a selectable option within This corroborates NovitaAI's listing of GLM-5 (modelKey "zai-org/glm-5") as a presently consumable model across third-party UIs and inference layers, reinforcing that the model is live in production rather than pre-release. While the entry is brief and lacks benchmark or pricing detail, it provides timely ecosystem evi

Videos about GLM-5

More models around GLM-5