Sulat.com
AI models
Get 10-25% off
Get 10-25% off from Qwen
Alibaba (China) logo

Model details

Qwen3 32B

Qwen3 32B is a 32-billion-parameter dense causal language model that belongs to the latest generation of the Qwen series, which also includes mixture-of-experts architectures. The model features 64 layers with grouped query attention (64 heads for queries, 8 for key-value pairs) and was designed from the ground up to support seamless switching between thinking mode for complex logical reasoning, mathematics, and code generation, and non-thinking mode for efficient general-purpose dialogue. This dual-mode capability allows the same model to handle deeply multi-step problem-solving alongside quick conversational responses without requiring separate specialist models. The architecture also emphasizes agent capabilities, enabling precise integration with external tools in both operational modes.

The model underwent extensive pretraining followed by post-training to develop its reasoning and instruction-following abilities, achieving performance that surpasses earlier QwQ thinking models and Qwen2.5 instruct models across mathematics, code generation, and commonsense reasoning benchmarks. Qwen3 32B demonstrates strong human preference alignment, excelling in creative writing, role-playing, and multi-turn dialogues. It supports over 100 languages and dialects with robust multilingual instruction-following and translation capabilities. Released under the Apache 2.0 license, it sits alongside a family of models ranging from 0.6B to 235B parameters, enabling flexibility from edge deployment to large-scale reasoning tasks. The model's open-weight availability and agentic tool-use proficiency make it well-suited for developers building autonomous workflows, research pipelines, and multilingual applications.

Alibaba (China)qwen3-32bqwen

Quick Info

Powered by
Provider
Alibaba (China)
Model key
qwen3-32b
Release date
Apr 1, 2025
Last updated
Apr 1, 2025
Knowledge cutoff
2025-04
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.287
Output token cost
$1.147

Limits

Output tokens
16,384 tokens
Context window
131,072 tokens

Transparent token rates

Compare Qwen3 32B pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Qwen3 32B

No articles yet. Fetch the latest news to show it here.

Videos about Qwen3 32B

More models around Qwen3 32B