Currently listed through these providers:
Model details
Qwen3 235B A22B Thinking (2507)
Qwen3-235B-A22B-Thinking-2507 is a Mixture-of-Experts causal language model built around 235 billion total parameters, activating 22 billion per forward pass across a dense expert routing system. Its architecture supports structured reasoning through a dedicated thinking mode that chains responses through explicit reasoning traces, designed specifically for tasks requiring methodical problem-solving rather than direct answering. With native context handling extending to over 262,000 tokens and Grouped Query Attention for efficient inference, the model targets high-complexity domains where logical depth, multi-step derivation, and extended context integration matter most.
The 2507 release marks the third iteration in a focused three-month push to deepen thinking capability, building on instruction-tuning and post-training enhancements to improve benchmark results among open-source thinking models. Beyond structured reasoning gains in mathematics, science, and coding, the update delivers measurable progress on general capabilities including tool usage, instruction following, and alignment with human preferences. Its open-weight availability under Apache 2.0 licensing makes it accessible for agents, automated workflows, and multilingual applications where step-by-step logical fidelity is essential.
Quick Info
Powered by- Provider
- DevPass (LLM Gateway)
- Model key
- qwen3-235b-a22b-thinking-2507
- Release date
- Jul 8, 2025
- Last updated
- Jul 8, 2025
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.30
- Output token cost
- $3.00
Limits
- Output tokens
- 8,192 tokens
- Context window
- 262,000 tokens
Transparent token rates
Compare Qwen3 235B A22B Thinking (2507) pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about Qwen3 235B A22B Thinking (2507)
No articles yet. Fetch the latest news to show it here.