Currently listed through these providers:
Model details
Qwen/Qwen3-235B-A22B-Thinking-2507
This model utilizes a Mixture-of-Experts architecture, featuring 235 billion total parameters with 22 billion parameters activated per forward pass. Built with 128 experts and 94 layers, it is specifically engineered as a thinking-only system that forces a structured reasoning process. This design intent focuses on delivering high-depth performance for tasks requiring human-level expertise, such as complex mathematics, scientific inquiry, coding, and logical deduction. By natively supporting a 262,144-token context window, the model is optimized for long-form generation and intricate, multi-step problem-solving scenarios.
The model underwent rigorous pre-training and post-training stages to refine its instruction-following and alignment with human preferences. Its development emphasizes agentic workflows and tool usage, allowing it to excel in environments that demand precise, step-by-step reasoning. By leveraging its specialized thinking mode, the model achieves state-of-the-art results among open-source alternatives, making it a robust choice for academic benchmarks and professional applications. Its architecture is well-suited for future-facing tasks that require deep analytical processing and reliable, structured output in challenging domains.
Quick Info
Powered by- Provider
- SiliconFlow (China)
- Model key
- Qwen/Qwen3-235B-A22B-Thinking-2507
- Release date
- Jul 28, 2025
- Last updated
- Nov 25, 2025
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.13
- Output token cost
- $0.60
Limits
- Output tokens
- 262,000 tokens
- Context window
- 262,000 tokens
Transparent token rates
Compare Qwen/Qwen3-235B-A22B-Thinking-2507 pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about Qwen/Qwen3-235B-A22B-Thinking-2507
No articles yet. Fetch the latest news to show it here.