Currently listed through these providers:
Model details
GLM 4.7
GLM-4.7 is the latest flagship in the GLM family from Z.ai (Zhipu AI), a developer focused on bringing frontier-level capability to open-source ecosystems. Built to handle complex multi-step development workflows, the model combines a large dense architecture with optimizations that target real-world agentic tasks, coding scenarios, and extended reasoning chains. Sources describe it as designed to handle multi-step tasks, with capabilities that match or surpass leading closed models in software engineering benchmarks, positioning the family as a credible open-weight alternative to proprietary giants. The model follows the well-received GLM-4.6 release, continuing Z.ai's trajectory of building models that prioritize practical developer workflows over chat-focused polish alone.
GLM-4.7 was released as open-source on Hugging Face, signaling Z.ai's commitment to making powerful development tools accessible without API lock-in. The training and post-training approach emphasized coding performance, agent execution, and reasoning depth, resulting in measurable gains that some independent evaluations cite as surpassing Gemini 3.0 Pro on specific benchmarks. A companion Flash variant uses a 30 billion parameter Mixture-of-Experts design that activates only a fraction of weights per token, enabling near-frontier performance on mid-range hardware while keeping inference costs and latency manageable. For teams seeking open-weight models that can handle intelligent agent pipelines, tool-augmented reasoning, and structured output scenarios, GLM-4.7 represents a concrete open-source option that does not require trading away architectural transparency for capability.
Quick Info
Powered by- Provider
- Venice AI
- Model key
- zai-org-glm-4.7
- Release date
- Dec 24, 2025
- Last updated
- Jun 11, 2026
- Knowledge cutoff
- 2025-04
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.55
- Output token cost
- $2.65
Limits
- Output tokens
- 16,384 tokens
- Context window
- 198,000 tokens
Transparent token rates
Compare GLM 4.7 pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about GLM 4.7
No articles yet. Fetch the latest news to show it here.