Currently listed through these providers:
Model details
Qwen3 235B A22B Instruct
The Qwen3-235B-A22B-Instruct-2507 is a Mixture of Experts causal language model that activates only 22 billion of its 235 billion total parameters during inference, making it computationally efficient while retaining broad capability coverage. Its MoE architecture draws from a pool of 128 experts, routing through 8 active experts per forward pass across 94 transformer layers with grouped query attention. This design enables the model to specialize across diverse domains—from mathematical reasoning and scientific understanding to code generation and multi-language comprehension—while maintaining a relatively lightweight inference footprint compared to dense models of similar total scale.
The July 2025 update represents a significant post-training iteration that brings substantial gains in instruction following, logical reasoning, text comprehension, and tool usage over the earlier Qwen3-235B checkpoint. The model was trained through both pre-training and post-training stages, and its non-thinking mode produces direct responses without generating intermediate reasoning blocks. Enhanced alignment with human preferences allows it to deliver higher-quality, more helpful responses in open-ended and subjective tasks. Its native 256K context window—with extension capability up to roughly 1 million tokens—makes it well suited for document analysis, long conversations, and complex multi-step reasoning where broad knowledge retrieval and extended context understanding are essential.
Quick Info
Powered by- Provider
- Abacus
- Model key
- Qwen/Qwen3-235B-A22B-Instruct-2507
- Release date
- Jul 1, 2025
- Last updated
- Jul 1, 2025
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.13
- Output token cost
- $0.60
Limits
- Output tokens
- 8,192 tokens
- Context window
- 262,144 tokens
Transparent token rates
Compare Qwen3 235B A22B Instruct pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about Qwen3 235B A22B Instruct
No articles yet. Fetch the latest news to show it here.