Sulat.com
AI models
Get 10-25% off
Get 10-25% off from Qwen
Alibaba logo

Model details

Qwen3.5 35B-A3B

Qwen3.5-35B-A3B is a sparse Mixture-of-Experts model that activates only 3 billion of its 35 billion total parameters per token, making it production-friendly without sacrificing capability. Its hybrid architecture combines Gated Delta Networks with linear attention mechanisms to deliver high-throughput inference with minimal latency and cost overhead. Early fusion training on multimodal tokens enables this model to reason across text, images, video, and audio with cross-generational parity to Qwen3, even outperforming dedicated Qwen3-VL models on reasoning, coding, agentic tasks, and visual understanding benchmarks. The architecture supports a native 262K token context window, extensible to 1 million tokens through YaRN extrapolation, positioning it for both long-document analysis and real-time multimodal interactions.

The model builds on Alibaba Cloud's reinforcement learning scaling work to achieve its benchmark results, representing a deliberate push toward models that deliver exceptional utility alongside architectural efficiency. Released as open weights with artifacts compatible with Hugging Face Transformers, vLLM, SGLang, and KTransformers, it is designed for developers who want the flexibility of self-hosting combined with strong out-of-the-box performance. Its tool-calling support and reasoning mode make it well-suited for agentic applications, while its multilingual capability across 201 languages broadens its practical reach. The combination of open access, efficiency, and multimodal reasoning positions this model as a versatile foundation for coding assistants, autonomous agents, and production pipelines that need vision-language capability without dense-model compute costs.

Alibabaqwen3.5-35b-a3bqwen

Quick Info

Powered by
Provider
Alibaba
Model key
qwen3.5-35b-a3b
Release date
Feb 23, 2026
Last updated
Feb 23, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.25
Output token cost
$2.00

Limits

Output tokens
65,536 tokens
Context window
262,144 tokens

Transparent token rates

Compare Qwen3.5 35B-A3B pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Qwen3.5 35B-A3B

Alibaba

CoverageRelease Notes

ThursdAI's February 2026 monthly releases roundup identifies Qwen 3.5 as a lead launch that month, explicitly listing the Medium Series open-source wave with a "35B / 3B active Qwen 3.5 Medium" configuration alongside the 27B dense and 122B-A10B variants. The page places Qwen3.5-35B-A3B within a broader wave of 16 open Among other February 2026 standouts cited in the same roundup are MiniMax M-2.5 at 80.2% SWE-Bench Verified, GLM-5 at 744B parameters, and Qwen3-Coder-Next at 70.6% SWE-Bench Verified, with the Qwen 3.5 Medium variants positioned as the open-weights cost-efficiency tier. The page catalogs 57 total releases with primary

Videos about Qwen3.5 35B-A3B

More models around Qwen3.5 35B-A3B