Currently listed through these providers:
Model details
Qwen Max
Qwen Max sits inside the broader Qwen family of large language models developed by the Qwen team at Alibaba, a lineage that has progressed through successive generations emphasizing stronger reasoning, coding, and long-horizon task handling. It is positioned as a text-only large language model, accepting text input and producing text output, making it a general-purpose choice for conversational, analytical, and content-generation workloads. Its placement within the Qwen ecosystem connects it to a wider research effort that has explored very large parameter counts, mixture-of-experts efficiency, and broader context windows in sibling and successor variants.
Because the supplied evidence covers a different, later Qwen-Max-class release rather than the 2024 catalog entry itself, this overview focuses on the role Qwen Max plays in practical deployments rather than on specific benchmark numbers. The model is well matched to text-centric assistants, document summarization, structured generation pipelines, and integration scenarios where temperature control and structured output formatting are useful. Teams evaluating it should weigh it against more recent Qwen variants for cutting-edge coding or agentic tasks, while considering Qwen Max as a stable, general-purpose option in the family for everyday language understanding and generation needs.
Quick Info
Powered by- Provider
- Ofox
- Model key
- qwen/qwen-max
- Release date
- Apr 3, 2024
- Last updated
- Jan 25, 2025
- Knowledge cutoff
- 2024-04
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.35
- Output token cost
- $1.38
Limits
- Output tokens
- 8,000 tokens
- Context window
- 32,000 tokens
Transparent token rates
Compare Qwen Max pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about Qwen Max
No articles yet. Fetch the latest news to show it here.