Currently listed through these providers:
Model details
Qwen3 32B
Qwen3 32B is a dense causal language model built with 32.8 billion parameters, structured across 64 layers to facilitate deep interaction between input features and non-linear functions. Designed as a core member of the Qwen3 series, the architecture features a unique capability to switch seamlessly between a dedicated thinking mode for complex logical, mathematical, and coding tasks and a non-thinking mode for efficient, general-purpose dialogue. This dual-mode design allows the model to maintain high performance across diverse scenarios, from creative writing and role-playing to precise, agent-based tool integration.
The model underwent extensive pre-training and post-training stages to achieve superior human preference alignment and multilingual proficiency across more than 100 languages and dialects. By leveraging these training methods, the model demonstrates significant advancements in instruction following and commonsense reasoning, positioning it as a strong performer for production-level workloads. Its design supports complex agentic tasks and deep reasoning, making it a practical choice for developers seeking a balance between high-level cognitive capabilities and efficient, scalable deployment in both enterprise and research environments.
Quick Info
Powered by- Provider
- Alibaba
- Model key
- qwen3-32b
- Release date
- Apr 1, 2025
- Last updated
- Apr 1, 2025
- Knowledge cutoff
- 2025-04
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.70
- Output token cost
- $2.80
Limits
- Output tokens
- 16,384 tokens
- Context window
- 131,072 tokens
Transparent token rates
Compare Qwen3 32B pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about Qwen3 32B
No articles yet. Fetch the latest news to show it here.