Currently listed through these providers:
Model details
Qwen2.5 32B Instruct
Qwen2.5 32B Instruct is a 32-billion-parameter language model in Alibaba's Qwen family of instruction-tuned models. It sits in the middle of the Qwen2.5 lineup, balancing capacity with practical deployability for users who want stronger reasoning than smaller Qwen variants while keeping inference costs manageable on accessible hardware. The instruct-tuned configuration means it is post-trained to follow natural-language directions, making it well suited for chat, drafting, summarization, question answering, code assistance, and structured tool use.
Released with open weights, Qwen2.5 32B Instruct can be self-hosted and adapted, which makes it attractive for teams that need privacy, customization, or integration into local pipelines. The model supports text-to-text tasks and includes features such as tool calling and temperature control, giving developers flexibility to shape response style and connect the model to external functions. Its very large context window lets it handle long documents, multi-turn conversations, and extended code bases, while the 32B parameter scale offers a strong mix of language understanding and generation quality for production use.
Quick Info
Powered by- Provider
- Alibaba
- Model key
- qwen2-5-32b-instruct
- Release date
- Sep 1, 2024
- Last updated
- Sep 1, 2024
- Knowledge cutoff
- 2024-04
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.70
- Output token cost
- $2.80
Limits
- Output tokens
- 8,192 tokens
- Context window
- 131,072 tokens
Transparent token rates
Compare Qwen2.5 32B Instruct pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about Qwen2.5 32B Instruct
No articles yet. Fetch the latest news to show it here.