Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
ModelScope logo

Model details

Qwen3 235B A22B Instruct 2507

Qwen3 235B A22B Instruct 2507 is a non-thinking variant in Alibaba's Qwen3 family, built as an alternative to the hybrid thinking-and-non-thinking Qwen3 235B release. It uses a Mixture-of Experts architecture with roughly 235 billion total parameters and about 22 billion activated per token, and the weights are distributed openly under an Apache-2.0 license on ModelScope in Transformers and Safetensors formats. Developers can run the model locally or host it through compatible inference providers, with the open-weight distribution making it a practical base for fine-tuning, research, and deployment that demands transparency around the model file itself.

This release targets users who want strong general capability without paying the latency cost of chain-of-thought reasoning. According to Alibaba's Qwen team and third-party coverage, it brings meaningful gains over the earlier Qwen3 235B hybrid model in instruction following, logical reasoning, mathematics, science, coding, and tool usage, along with broader long-tail knowledge coverage and better alignment on subjective and open-ended tasks. It is described as achieving state-of-the-art results among non-reasoning models on the Artificial Analysis Intelligence Index, a blended benchmark across general knowledge, reasoning, coding, and STEM, outperforming frontier peers in that comparison. That profile makes it well suited to production assistants, agentic workflows, and multilingual applications where a fast, instruction-tuned base model is more useful than a slower reasoning variant.

ModelScopeQwen/Qwen3-235B-A22B-Instruct-2507qwen

Quick Info

Powered by
Provider
ModelScope
Model key
Qwen/Qwen3-235B-A22B-Instruct-2507
Release date
Apr 28, 2025
Last updated
Jul 21, 2025
Knowledge cutoff
2025-04
Input modalities
Output modalities
Capabilities

Cost

A provider subscription or plan supersedes token-based pricing for this model.

Limits

Output tokens
131,072 tokens
Context window
262,144 tokens

Latest news about Qwen3 235B A22B Instruct 2507

Nebius Token Factory

Coverage

The ModelScope model card confirms Qwen/Qwen3-235B-A22B-Instruct-2507 as a first-party Alibaba release: a Text Generation model with 235.09 billion parameters in Transformers, Safetensors, and PyTorch formats, tagged qwen3 / moe, and distributed under the Apache 2.0 license. The listing has accumulated over 300,886 dow The artifact page records the model as last updated on September 17, 2025, providing authoritative provenance that the 235B-A22B-Instruct-2507 variant is the canonical Qwen release rather than a serving-provider-specific build. It serves as a primary source anchor verifying the model's official identity, licensing, and

Videos about Qwen3 235B A22B Instruct 2507

More models around Qwen3 235B A22B Instruct 2507