Currently listed through these providers:
Model details
GLM-4.6V
The GLM-4.6V series represents a significant step in multimodal architecture, utilizing a Mixture-of-Experts design to balance high-level reasoning with operational efficiency. The flagship 106B model is engineered for demanding cloud and high-performance cluster environments, while the 9B Flash version provides a lightweight alternative for local, low-latency tasks. By supporting multi-resolution image processing up to 4K, these models are built to handle complex visual inputs, such as screenshots and document pages, with high precision. This design intent focuses on creating a unified foundation where visual perception directly informs executable actions, making the series particularly effective for agentic workflows.
Building upon the technical lineage of the GLM-V series, these models introduce native function calling to eliminate the information loss often associated with traditional text-based conversions. This capability allows the models to interact directly with external tools for tasks like image searching, chart recognition, and front-end automation. By integrating these features into a cohesive training system, the series achieves state-of-the-art performance in visual understanding among models of similar scale. With its focus on bridging the gap between visual analysis and real-world utility, the series offers a robust framework for developers looking to build sophisticated multimodal agents that can operate across diverse business and technical scenarios.
Quick Info
Powered by- Provider
- AIHubMix
- Model key
- glm-4.6v
- Release date
- Dec 8, 2025
- Last updated
- Dec 8, 2025
- Knowledge cutoff
- 2025-04
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.137
- Output token cost
- $0.411
Limits
- Output tokens
- 32,768 tokens
- Context window
- 128,000 tokens
Transparent token rates
Compare GLM-4.6V pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.