Currently listed through these providers:
Model details
GLM-4.7
GLM-4.7 is built on a Mixture-of-Experts (MoE) Transformer architecture that activates only a fraction of its total parameters during inference, enabling computational efficiency without sacrificing depth. The model supports a 200,000-token context window paired with a 128,000-token output capacity—specifications that let it generate entire software frameworks or sustain complex multi-step reasoning chains in a single pass. Designed specifically for real-world development workflows, it prioritizes agentic tasks and coding scenarios where sustained coherence and large-scale code generation matter more than simple chat interactions.
Released in late December 2025 as the latest flagship in the GLM family, GLM-4.7 introduces meaningful upgrades to programming capability and multi-step execution stability. The model's open-weight availability on Hugging Face signals Zhipu's commitment to transparency and community-driven refinement, positioning it for adoption among developers who want frontier-level performance without proprietary lock-in. It performs competitively across academic, finance, legal, and programming benchmarks, and its emphasis on agentic workflow stability makes it particularly suitable for intelligent assistants and autonomous coding agents operating in production environments.
Quick Info
Powered by- Provider
- Zhipu AI
- Model key
- glm-4.7
- Release date
- Dec 22, 2025
- Last updated
- Dec 22, 2025
- Knowledge cutoff
- 2025-04
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.60
- Output token cost
- $2.20
Limits
- Output tokens
- 131,072 tokens
- Context window
- 204,800 tokens
Transparent token rates
Compare GLM-4.7 pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about GLM-4.7
No articles yet. Fetch the latest news to show it here.