Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Vertex logo

Model details

Qwen3 235B A22B Instruct

Qwen3 235B A22B Instruct is a mixture-of-experts language model from the Qwen3 family, designed to balance broad capability with efficient inference. Its MoE architecture activates 22 billion parameters per forward pass while retaining access to the full 235 billion parameter space, giving it the knowledge breadth of a much larger dense model without proportional serving costs. The model was built with extensive pretraining and covers 119 languages and dialects, making it well-suited for multilingual applications. Benchmarks cited in its evaluation include AIME and HMMT for mathematics, Arena-Hard and WritingBench for alignment and open-ended tasks, where the instruct-tuned variant shows gains over its base counterpart in knowledge coverage, long-context reasoning, and coding performance.

This instruction-tuned version builds on the Qwen3-235B architecture with explicit optimization for instruction following, logical reasoning, and tool calling. The model does not implement a thinking mode by default, prioritizing immediate responses for latency-sensitive tasks, though the broader Qwen3 family offers hybrid reasoning options for scenarios requiring extended step-by-step computation. Its support for tool calling and the Model Context Protocol makes it a natural fit for agentic pipelines and multi-step workflows where the model orchestrates external tools or APIs. For developers building enterprise applications, multilingual content systems, or research tools, the combination of open weights, strong benchmark standing against comparable open-source models, and native long-context support positions it as a practical choice for production deployments that need both capability and operational flexibility.

Vertexqwen/qwen3-235b-a22b-instruct-2507-maasqwendeprecated

Quick Info

Powered by
Provider
Vertex
Model key
qwen/qwen3-235b-a22b-instruct-2507-maas
Release date
Aug 13, 2025
Last updated
Aug 13, 2025
AI SDK package
@ai-sdk/openai-compatible
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.22
Output token cost
$0.88

Limits

Output tokens
16,384 tokens
Context window
262,144 tokens

Transparent token rates

Compare Qwen3 235B A22B Instruct pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Qwen3 235B A22B Instruct

Vertex

Coverage

Alibaba's Qwen team has published the official model card for Qwen3-235B-A22B-Instruct-2507 on Hugging Face, describing it as the updated non-thinking variant of the Qwen3-235B-A22B family. The model is a Mixture-of-Experts causal language model with 235B total parameters, 22B activated per inference, 94 layers, 64 Q-a The release highlights listed on the card include significant gains in instruction following, logical reasoning, text comprehension, mathematics, science, coding, and tool usage; broader long-tail multilingual knowledge coverage; better alignment on subjective and open-ended tasks; and enhanced 256K long-context unders

Videos about Qwen3 235B A22B Instruct

More models around Qwen3 235B A22B Instruct