Currently listed through these providers:
Model details
GLM-4.7
GLM-4.7 is a Mixture of Experts foundation model designed to unify reasoning, coding, and agentic capabilities within a single architecture. With hundreds of billions of total parameters but a much smaller active parameter count during inference, the model balances computational efficiency with strong performance across demanding tasks. It targets developers and enterprises working in production environments where tasks span long cycles, require frequent tool interactions, and demand consistent behavior across multiple steps. The architecture supports thinking before acting, enabling the model to deliberate on complex problems rather than immediately responding, which proves valuable for terminal-based tasks, multilingual coding scenarios, and framework integrations like Claude Code, Kilo Code, Cline, and Roo Code.
The model builds on its predecessor with targeted improvements in agentic coding, mathematical reasoning, and tool use. Benchmark gains are substantial: double-digit percentage improvements on SWE-bench for software engineering tasks, Terminal Bench for command-line workflows, and the HLE benchmark for advanced mathematical reasoning. GLM-4.7 also introduces "Vibe Coding" capabilities, producing cleaner, more modern webpage layouts and better-formatted slides with improved sizing accuracy. Tool calling follows the OpenAI-style format, and the model demonstrates significantly better performance on τ²-Bench and web browsing benchmarks like BrowseComp. As an open-weight model, it serves developers seeking a production-ready foundation that can handle intelligent agent applications while remaining accessible for customization and self-hosting through frameworks like vLLM and SGLang.
Quick Info
Powered by- Provider
- Alibaba Coding Plan
- Model key
- glm-4.7
- Release date
- Dec 22, 2025
- Last updated
- Dec 22, 2025
- Knowledge cutoff
- 2025-04
- Input modalities
- Output modalities
- Capabilities
Cost
A provider subscription or plan supersedes token-based pricing for this model.
Limits
- Output tokens
- 16,384 tokens
- Context window
- 202,752 tokens
Latest news about GLM-4.7
No articles yet. Fetch the latest news to show it here.