Currently listed through these providers:
Model details
GLM 4.7 Flash (Free)
GLM 4.7 Flash is built on a Mixture of Experts architecture with approximately 30 billion parameters and an A3B activation pattern, meaning only a subset of the model activates during each forward pass. This design choice trades off raw depth for computational efficiency, making the model particularly well-suited for developer workflows that demand quick turnarounds. The architecture targets local coding tasks and agent-based applications where speed and responsiveness matter more than exhaustive reasoning.
The model serves as a practical entry point for developers handling straightforward coding tasks, document formatting, and quick lookups. While it deliberately sacrifices depth to prioritize speed, it holds its own on simple completions and formatting work. Complex debugging, multi-file refactoring, and sophisticated agentic workflows remain areas where premium models outperform it, but GLM 4.7 Flash provides a solid free option for rapid prototyping and lightweight tasks where cost matters more than exhaustive analysis.
Quick Info
Powered by- Provider
- ZenMux
- Model key
- z-ai/glm-4.7-flash-free
- Release date
- Jan 19, 2026
- Last updated
- Jan 19, 2026
- Knowledge cutoff
- 2025-04
- Input modalities
- Output modalities
- Capabilities
Cost
A provider subscription or plan supersedes token-based pricing for this model.
Limits
- Output tokens
- 128,000 tokens
- Context window
- 200,000 tokens
Latest news about GLM 4.7 Flash (Free)
No articles yet. Fetch the latest news to show it here.
Videos about GLM 4.7 Flash (Free)
More models around GLM 4.7 Flash (Free)
This exact model name is also listed by 6 other providers.