Currently listed through these providers:
Model details
GLM-4 32B (0414-128k)
GLM-4-32B is an instruction-tuned language model with 32 billion parameters that punches above its weight class. Built by the same research lab behind the thudm models, it was pre-trained on 15 terabytes of high-quality data to deliver performance competitive with much larger mainstream models. The architecture is designed to be highly cost-effective while maintaining strong capabilities across complex reasoning and generation tasks. With a 128,000-token context window, it can handle lengthy documents and multi-turn conversations that would overwhelm smaller context models.
The model particularly excels in engineering code generation, artifact creation, function calling, search-based question answering, and report writing. Its training emphasis on tool use and intelligent task execution makes it well-suited for applications requiring precise function calling and structured outputs. The combination of strong code-related abilities, substantial context capacity, and instruction-following refinement positions this 32B model as a practical choice for developers seeking large-language-model capabilities without the operational overhead of larger deployments. The model was released with a knowledge cutoff around April 2025, ensuring reasonable currency for general knowledge queries.
Quick Info
Powered by- Provider
- DevPass (LLM Gateway)
- Model key
- glm-4-32b-0414-128k
- Release date
- Apr 14, 2025
- Last updated
- Apr 14, 2025
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.10
- Output token cost
- $0.10
Limits
- Output tokens
- 16,384 tokens
- Context window
- 128,000 tokens
Transparent token rates
Compare GLM-4 32B (0414-128k) pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about GLM-4 32B (0414-128k)
No articles yet. Fetch the latest news to show it here.