NovitaAI
Discover more about what's new at AWS with Minimax M2.5 and GLM 5 models now available on Amazon Bedrock
Model details
GLM-5 is a large open-weights mixture-of-experts model positioned for complex systems engineering and long-horizon agentic work rather than casual code generation. The architecture scales from the prior generation's 355B parameters with 32B active to 744B parameters with 40B active, supported by a pre-training corpus expanded from 23T to 28.5T tokens. To keep deployment economical at this scale, GLM-5 integrates DeepSeek Sparse Attention, which compresses long-context computation without sacrificing the model's ability to handle extended inputs. Together, these changes mark a deliberate step from snippet-level code assistance toward end-to-end engineering, where the model is expected to plan, coordinate tools, and sustain multi-step tasks across long contexts.
On the post-training side, the team paired a new asynchronous reinforcement-learning infrastructure called slime with continued scale, enabling more granular iterations on long-horizon task environments. The result is broad gains on academic benchmarks relative to earlier GLM variants, with the model also surfaced through developer-friendly channels such as Novita AI's Model APIs and Amazon Bedrock, confirming production availability for teams that want to wire it into agentic pipelines. Practical fit is strongest for organizations building software agents, automation backends, or research prototypes that need a capable, open-weights base they can self-host or route through managed inference, and who value long-context reasoning paired with sparse-attention efficiency over a smaller, denser alternative.
Transparent token rates
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
NovitaAI
Discover more about what's new at AWS with Minimax M2.5 and GLM 5 models now available on Amazon Bedrock
NovitaAI
SINGAPORE, February 16, 2026--GLM-5, newly released as open source, signals a broader shift in artificial intelligence. Large language models are moving beyond generating code snippets or interface prototypes toward building complete systems and carrying out complex, end-to-end tasks. The change marks a transition from
NovitaAI
Build with GLM-5 on Novita AI. 744B MoE architecture for expert agentic coding. Claim your free credits and start building today!, Post a Comment
NovitaAI
Big-AGI's public changelog lists "Z.ai GLM-5" support added alongside Moonshot Kimi K2.7-code on June 16, 2026, confirming that the Z.ai GLM-5 model is generally reachable through the OpenRouter-style routing that consumer apps aggregate. The changelog entry indicates the model is exposed as a selectable option within This corroborates NovitaAI's listing of GLM-5 (modelKey "zai-org/glm-5") as a presently consumable model across third-party UIs and inference layers, reinforcing that the model is live in production rather than pre-release. While the entry is brief and lacks benchmark or pricing detail, it provides timely ecosystem evi