Currently listed through these providers:
Model details
Qwen3 235B A22B Thinking 2507
The Qwen3 235B A22B Thinking 2507 is a causal language model built on a Mixture-of-Experts architecture, utilizing 128 total experts with 8 active at any given time to balance efficiency and power. Designed specifically for deep cognitive tasks, the model operates in a dedicated thinking mode that enforces structured, step-by-step logical processing. This design intent makes it particularly effective for challenging domains such as mathematics, science, and complex coding, where it aims to provide the depth of reasoning typically associated with human expertise.
Developed through extensive pre-training and post-training stages, this model represents a significant advancement in open-source reasoning capabilities. It has been refined to excel in instruction following and tool usage, making it a robust choice for agentic workflows that require reliable, multi-step output. By leveraging its large parameter count and native support for extended context, the model achieves state-of-the-art performance on various academic benchmarks, positioning it as a primary tool for users who need to solve intricate problems that demand both breadth of knowledge and rigorous analytical depth.
Quick Info
Powered by- Provider
- OpenRouter
- Model key
- qwen/qwen3-235b-a22b-thinking-2507
- Release date
- Jul 25, 2025
- Last updated
- Jul 25, 2025
- Knowledge cutoff
- 2025-06-30
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.23
- Output token cost
- $2.30
Limits
- Output tokens
- 117,964 tokens
- Context window
- 131,072 tokens
Transparent token rates
Compare Qwen3 235B A22B Thinking 2507 pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about Qwen3 235B A22B Thinking 2507
No articles yet. Fetch the latest news to show it here.