Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Abacus logo

Model details

Qwen3 235B A22B Instruct

The Qwen3-235B-A22B-Instruct-2507 is a Mixture of Experts causal language model that activates only 22 billion of its 235 billion total parameters during inference, making it computationally efficient while retaining broad capability coverage. Its MoE architecture draws from a pool of 128 experts, routing through 8 active experts per forward pass across 94 transformer layers with grouped query attention. This design enables the model to specialize across diverse domains—from mathematical reasoning and scientific understanding to code generation and multi-language comprehension—while maintaining a relatively lightweight inference footprint compared to dense models of similar total scale.

The July 2025 update represents a significant post-training iteration that brings substantial gains in instruction following, logical reasoning, text comprehension, and tool usage over the earlier Qwen3-235B checkpoint. The model was trained through both pre-training and post-training stages, and its non-thinking mode produces direct responses without generating intermediate reasoning blocks. Enhanced alignment with human preferences allows it to deliver higher-quality, more helpful responses in open-ended and subjective tasks. Its native 256K context window—with extension capability up to roughly 1 million tokens—makes it well suited for document analysis, long conversations, and complex multi-step reasoning where broad knowledge retrieval and extended context understanding are essential.

AbacusQwen/Qwen3-235B-A22B-Instruct-2507qwen

Quick Info

Powered by
Provider
Abacus
Model key
Qwen/Qwen3-235B-A22B-Instruct-2507
Release date
Jul 1, 2025
Last updated
Jul 1, 2025
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.13
Output token cost
$0.60

Limits

Output tokens
8,192 tokens
Context window
262,144 tokens

Transparent token rates

Compare Qwen3 235B A22B Instruct pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Qwen3 235B A22B Instruct

No articles yet. Fetch the latest news to show it here.

Videos about Qwen3 235B A22B Instruct

More models around Qwen3 235B A22B Instruct