Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Vercel AI Gateway logo

Model details

Qwen3 235B A22B Thinking 2507

Qwen3-235B-A22B-Thinking-2507 is a Mixture-of-Experts language model with 235 billion total parameters distributed across 128 experts, activating 22 billion per forward pass to balance massive capacity against practical compute budgets. Its architecture draws from 94 transformer layers with grouped-query attention, delivering strong reasoning quality within a manageable footprint. This thinking-only variant enforces structured step-by-step reasoning before generating final outputs, making it especially effective for tasks that demand precision across logical reasoning, mathematics, science, and software engineering. The model carries an enhanced 256K-token context window natively, allowing it to process extended documents or maintain coherence across long conversation histories without fragmenting understanding.

The 2507 release represents a deliberate refinement of earlier Qwen3-235B-A22B checkpoints, incorporating three months of targeted improvements to both the depth and breadth of the model's reasoning capabilities. Instruction-tuning shapes the model's ability to follow structured guidance, execute tool calls, and operate effectively within agentic pipelines, while alignment refinements strengthen outputs for real-world deployment scenarios. The model demonstrates state-of-the-art results among open-source thinking models on academic and coding benchmarks, surpassing many closed alternatives in structured reasoning tasks. With a default reasoning mode baked into its chat template and support for high-token outputs up to 81,920 tokens, this model thrives in scenarios where complex problems require room to think and iterate before arriving at answers.

Vercel AI Gatewayalibaba/qwen3-235b-a22b-thinkingqwen

Quick Info

Powered by
Provider
Vercel AI Gateway
Model key
alibaba/qwen3-235b-a22b-thinking
Release date
Sep 23, 2025
Last updated
Apr 1, 2025
Knowledge cutoff
2025-04
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.40
Output token cost
$4.00

Limits

Output tokens
32,768 tokens
Context window
131,072 tokens

Transparent token rates

Compare Qwen3 235B A22B Thinking 2507 pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Qwen3 235B A22B Thinking 2507

No articles yet. Fetch the latest news to show it here.

Videos about Qwen3 235B A22B Thinking 2507

More models around Qwen3 235B A22B Thinking 2507