Currently listed through these providers:
Model details
Qwen3 32B
Qwen3 32B is a dense causal language model built with 32.8 billion parameters, structured across 64 layers to facilitate deep interaction between input features. Designed as a versatile powerhouse, it features a unique architecture that allows for seamless switching between a specialized thinking mode for complex logical, mathematical, and coding tasks and a non-thinking mode for efficient, general-purpose dialogue. This dual-mode capability ensures the model maintains high performance across a wide spectrum of use cases, from creative writing and role-playing to rigorous analytical problem-solving.
The model benefits from extensive pre-training and post-training stages, resulting in superior human preference alignment and robust instruction-following capabilities. It demonstrates significant advancements in agentic workflows, enabling precise integration with external tools to handle complex, multi-step tasks. With native support for over 100 languages and dialects, the model is well-suited for global applications requiring high-quality translation and multilingual interaction. Its design emphasizes both efficiency and depth, making it a strong candidate for developers building production-level workloads that demand reliable reasoning and natural, immersive conversational experiences.
Quick Info
Powered by- Provider
- DevPass (LLM Gateway)
- Model key
- qwen3-32b
- Release date
- Apr 1, 2025
- Last updated
- Apr 1, 2025
- Knowledge cutoff
- 2025-04
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.36
- Output token cost
- $0.87
Limits
- Output tokens
- 16,384 tokens
- Context window
- 40,960 tokens
Transparent token rates
Compare Qwen3 32B pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about Qwen3 32B
No articles yet. Fetch the latest news to show it here.