Currently listed through these providers:
Model details
Qwen: Qwen3 Max Thinking (retires Oct 9)
Qwen3 Max Thinking serves as the flagship reasoning model within the Qwen3 series, engineered specifically to handle high-stakes cognitive tasks that demand deep, multi-step analysis. By significantly scaling model capacity, it is designed to excel in scenarios requiring high factual accuracy, complex logical deduction, and precise instruction following. Its architecture is built to support sophisticated agentic behavior, allowing it to function effectively in environments where nuanced decision-making and reliable output are essential.
The model benefits from extensive reinforcement learning compute, which enhances its alignment with human preferences and overall performance. It incorporates multiple technical innovations, including adaptive tool-use and advanced test-time scaling techniques, to improve its problem-solving reliability. These advancements position the model as a strong candidate for demanding workflows that require both reasoning depth and the ability to interact with external tools, making it a versatile asset for users seeking high-performance language processing.
Quick Info
Powered by- Provider
- Kilo Gateway
- Model key
- qwen/qwen3-max-thinking
- Release date
- Feb 9, 2026
- Last updated
- Feb 9, 2026
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.78
- Output token cost
- $3.90
Limits
- Output tokens
- 65,536 tokens
- Context window
- 262,144 tokens
Transparent token rates
Compare Qwen: Qwen3 Max Thinking (retires Oct 9) pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.