Currently listed through these providers:
Model details
Qwen3 Max
Qwen3 Max belongs to the Qwen family of large language models developed by Alibaba Cloud's Qwen team, who also publish the model's reasoning-focused sibling Qwen3-Max-Thinking. The team's official blog positions the Thinking variant as a flagship reasoning model built by scaling model parameters and applying substantial reinforcement learning compute, with stated improvements across factual knowledge, complex reasoning, instruction following, alignment with human preferences, and agent capabilities. This indicates that the broader Qwen3-Max line is oriented toward deliberate, multi-step problem solving rather than purely latency-sensitive inference, making it a practical fit for analytical workflows that benefit from deeper deliberation.
Within that family, Qwen3-Max-Thinking is reported on the Qwen blog to perform comparably to leading frontier reasoning systems on a set of 19 established benchmarks, and to surpass specific reasoning competitors on key tests through advanced test-time scaling techniques. The variant additionally introduces adaptive tool-use capabilities that allow on-demand retrieval and code interpreter invocation, signaling that the Qwen3 Max line is designed not just for static question answering but for agent-style tasks that mix reasoning with external actions. For practitioners, this translates into a model family well suited to knowledge-intensive assistants, multi-step analytical pipelines, and tool-augmented applications where careful reasoning is more valuable than raw response speed.
Quick Info
Powered by- Provider
- Ofox
- Model key
- qwen/qwen3-max
- Release date
- Sep 23, 2025
- Last updated
- Sep 23, 2025
- Knowledge cutoff
- 2025-04
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.36
- Output token cost
- $1.43
Limits
- Output tokens
- 64,000 tokens
- Context window
- 256,000 tokens
Transparent token rates
Compare Qwen3 Max pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about Qwen3 Max
No articles yet. Fetch the latest news to show it here.