Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
302.AI logo

Model details

GLM-4.5V

GLM-4.5V is a vision-language model built upon ZhipuAI's flagship GLM-4.5-Air text foundation model, featuring 106 billion total parameters with 12 billion active parameters during inference. The model continues the technical approach established by GLM-4.1V-Thinking, designed specifically to move beyond basic multimodal perception toward enhanced reasoning capabilities. This architecture enables complex problem solving, long-context understanding, and the development of capable multimodal agents, positioning it as a versatile platform for demanding visual-language tasks.

Training leverages scalable reinforcement learning techniques to cultivate versatile multimodal reasoning abilities, which contributed to the model achieving state-of-the-art performance among similarly scaled models on 42 public vision-language benchmarks. As an open-source release from the GLM-V team, GLM-4.5V supports both thinking and non-thinking modes, allowing developers to trade off depth and speed depending on task requirements. The model is available through Hugging Face with a desktop assistant demo application and an online chat interface, making it accessible for both research exploration and practical application development.

302.AIglm-4.5vglm

Quick Info

Powered by
Provider
302.AI
Model key
glm-4.5v
Release date
Aug 12, 2025
Last updated
Aug 12, 2025
Knowledge cutoff
2025-04
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.29
Output token cost
$0.86

Limits

Output tokens
16,384 tokens
Context window
64,000 tokens

Transparent token rates

Compare GLM-4.5V pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about GLM-4.5V

No articles yet. Fetch the latest news to show it here.

Videos about GLM-4.5V

More models around GLM-4.5V