Currently listed through these providers:
Model details
Qwen3 235B A22B Thinking 2507
Qwen3-235B-A22B-Thinking-2507 is a Mixture-of-Experts language model with 235 billion total parameters distributed across 128 experts, activating 22 billion per forward pass to balance massive capacity against practical compute budgets. Its architecture draws from 94 transformer layers with grouped-query attention, delivering strong reasoning quality within a manageable footprint. This thinking-only variant enforces structured step-by-step reasoning before generating final outputs, making it especially effective for tasks that demand precision across logical reasoning, mathematics, science, and software engineering. The model carries an enhanced 256K-token context window natively, allowing it to process extended documents or maintain coherence across long conversation histories without fragmenting understanding.
The 2507 release represents a deliberate refinement of earlier Qwen3-235B-A22B checkpoints, incorporating three months of targeted improvements to both the depth and breadth of the model's reasoning capabilities. Instruction-tuning shapes the model's ability to follow structured guidance, execute tool calls, and operate effectively within agentic pipelines, while alignment refinements strengthen outputs for real-world deployment scenarios. The model demonstrates state-of-the-art results among open-source thinking models on academic and coding benchmarks, surpassing many closed alternatives in structured reasoning tasks. With a default reasoning mode baked into its chat template and support for high-token outputs up to 81,920 tokens, this model thrives in scenarios where complex problems require room to think and iterate before arriving at answers.
Quick Info
Powered by- Provider
- Vercel AI Gateway
- Model key
- alibaba/qwen3-235b-a22b-thinking
- Release date
- Sep 23, 2025
- Last updated
- Apr 1, 2025
- Knowledge cutoff
- 2025-04
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.40
- Output token cost
- $4.00
Limits
- Output tokens
- 32,768 tokens
- Context window
- 131,072 tokens
Transparent token rates
Compare Qwen3 235B A22B Thinking 2507 pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about Qwen3 235B A22B Thinking 2507
No articles yet. Fetch the latest news to show it here.