Currently listed through these providers:
Model details
Qwen3 32B
Qwen3 32B is a 32-billion-parameter dense causal language model that belongs to the latest generation of the Qwen series, which also includes mixture-of-experts architectures. The model features 64 layers with grouped query attention (64 heads for queries, 8 for key-value pairs) and was designed from the ground up to support seamless switching between thinking mode for complex logical reasoning, mathematics, and code generation, and non-thinking mode for efficient general-purpose dialogue. This dual-mode capability allows the same model to handle deeply multi-step problem-solving alongside quick conversational responses without requiring separate specialist models. The architecture also emphasizes agent capabilities, enabling precise integration with external tools in both operational modes.
The model underwent extensive pretraining followed by post-training to develop its reasoning and instruction-following abilities, achieving performance that surpasses earlier QwQ thinking models and Qwen2.5 instruct models across mathematics, code generation, and commonsense reasoning benchmarks. Qwen3 32B demonstrates strong human preference alignment, excelling in creative writing, role-playing, and multi-turn dialogues. It supports over 100 languages and dialects with robust multilingual instruction-following and translation capabilities. Released under the Apache 2.0 license, it sits alongside a family of models ranging from 0.6B to 235B parameters, enabling flexibility from edge deployment to large-scale reasoning tasks. The model's open-weight availability and agentic tool-use proficiency make it well-suited for developers building autonomous workflows, research pipelines, and multilingual applications.
Quick Info
Powered by- Provider
- Alibaba (China)
- Model key
- qwen3-32b
- Release date
- Apr 1, 2025
- Last updated
- Apr 1, 2025
- Knowledge cutoff
- 2025-04
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.287
- Output token cost
- $1.147
Limits
- Output tokens
- 16,384 tokens
- Context window
- 131,072 tokens
Transparent token rates
Compare Qwen3 32B pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about Qwen3 32B
No articles yet. Fetch the latest news to show it here.