Currently listed through these providers:
Model details
GLM-5.1
GLM-5.1 is positioned as a next-generation flagship model built specifically for agentic engineering workflows, with its developers claiming significantly stronger coding capabilities than its predecessor and reported state-of-the-art performance on SWE-Bench Pro, where it leads the earlier GLM-5 by a wide margin. The model is published with open weights under the MIT license, making it accessible for self-hosted deployments and third-party inference platforms. Its large scale supports extended reasoning across complex, multi-step software engineering tasks rather than short, single-turn interactions.
According to the Ollama library listing, GLM-5.1 carries 756 billion parameters and operates within a 198K-token context window, giving it substantial room to ingest large codebases, documentation, and multi-file projects in a single pass. The model is tagged with "thinking" capability and exposes "tool calls" support, indicating it is designed to participate in agent loops that require planning, reflection, and invocation of external functions. These characteristics make GLM-5.1 a practical fit for developers building autonomous coding agents, repository-level refactoring tools, and AI-assisted debugging pipelines where deep contextual understanding and structured reasoning across long inputs matter most.
Quick Info
Powered by- Provider
- DigitalOcean
- Model key
- glm-5.1
- Release date
- Apr 7, 2026
- Last updated
- Apr 7, 2026
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $1.30
- Output token cost
- $4.30
Limits
- Output tokens
- 163,840 tokens
- Context window
- 163,840 tokens