Currently listed through these providers:
Model details
Qwen Max
Qwen Max is a large-scale mixture-of-experts model engineered to deliver high-level inference performance, particularly for intricate, multi-step tasks. Designed as a versatile solution for demanding workflows, the architecture has evolved to incorporate specialized upgrades in agent programming and tool invocation. This focus on functional capability ensures the model is well-suited for environments requiring reliable interaction with external tools and complex logic, making it a robust choice for developers building sophisticated automated systems.
The model lineage is built upon a foundation of pre-training on over 20 trillion tokens, followed by rigorous post-training processes including curated supervised fine-tuning and reinforcement learning from human feedback. This combination of massive-scale training and targeted alignment methodologies enables the model to maintain high performance across diverse applications. With its ongoing refinements in agentic behavior and tool-use efficiency, the model is positioned as a forward-looking tool for developers seeking to deploy scalable, high-reasoning AI applications in production environments.
Quick Info
Powered by- Provider
- DevPass (LLM Gateway)
- Model key
- qwen-max
- Release date
- Apr 3, 2024
- Last updated
- Jan 25, 2025
- Knowledge cutoff
- 2024-04
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $1.60
- Output token cost
- $6.40
Limits
- Output tokens
- 8,192 tokens
- Context window
- 32,768 tokens
Transparent token rates
Compare Qwen Max pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about Qwen Max
No articles yet. Fetch the latest news to show it here.