Currently listed through these providers:
Model details
Qwen/Qwen3-32B
Qwen3-32B is Alibaba Cloud's flagship dense large language model in the Qwen3 family, built with 32.8 billion parameters across 64 transformer layers and grouped query attention (64 query heads, 8 key-value heads) to balance capability with deployment accessibility. The model uniquely supports seamless switching between thinking mode for deep analytical tasks like math, coding, and logical reasoning and non-thinking mode for rapid general-purpose responses—all within a single model. Its architecture emphasizes agent capabilities, enabling precise integration with external tools and achieving leading performance among open-source models on complex agent-based tasks.
Trained on an extensive corpus of 36 trillion tokens across 119 languages and dialects, Qwen3-32B brings strong multilingual fluency, instruction-following, and human preference alignment to a wide range of conversational and creative scenarios. The model excels in creative writing, role-playing, and multi-turn dialogues, while its agent tooling support makes it well-suited for building interactive AI applications. Released under the Apache 2.0 license, it offers both researchers and commercial developers unrestricted access to cutting-edge AI capabilities, positioning it as a competitive alternative to proprietary models in many practical deployments.
Quick Info
Powered by- Provider
- SiliconFlow (China)
- Model key
- Qwen/Qwen3-32B
- Release date
- Apr 30, 2025
- Last updated
- Nov 25, 2025
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.14
- Output token cost
- $0.57
Limits
- Output tokens
- 131,000 tokens
- Context window
- 131,000 tokens
Transparent token rates
Compare Qwen/Qwen3-32B pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about Qwen/Qwen3-32B
No articles yet. Fetch the latest news to show it here.