Currently listed through these providers:
Model details
GLM 4.6
GLM 4.6 is a coding-focused large language model developed by Z.AI (Zhipu AI) that builds on the GLM-4 and 4.5 lineage with targeted improvements in reasoning performance, tool integration, and deployment efficiency. The model is positioned for developers, researchers, and enterprises looking for an open-access alternative to leading closed systems, and it arrives alongside both an official Z.AI launch and a Hugging Face release. Its design emphasizes practical programming tasks and agentic workflows, reflecting Zhipu's broader ambition to make the GLM family a global competitor to frontier models from OpenAI, Anthropic, and Google.
A defining practical feature is the expanded 204.8K token context window, which lets GLM 4.6 ingest and reason over large codebases and long agent transcripts without aggressive truncation. The model ships with reasoning and tool use capabilities, enabling structured multi-step problem solving and reliable function calling for agent pipelines. Combined with its open-weight availability, these traits make GLM 4.6 a strong fit for code-heavy applications such as repository-scale code review, autonomous coding agents, and enterprise workflows that need transparency, customization, and long-context handling beyond what typical coding assistants offer.
Quick Info
Powered by- Provider
- Vercel AI Gateway
- Model key
- zai/glm-4.6
- Release date
- Sep 30, 2025
- Last updated
- Sep 30, 2025
- Knowledge cutoff
- 2025-04
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.60
- Output token cost
- $2.20
Limits
- Output tokens
- 96,000 tokens
- Context window
- 200,000 tokens
Transparent token rates
Compare GLM 4.6 pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.