Sulat.com
AI models
Get 10-25% off
Get 10-25% off from Qwen
Alibaba logo

Model details

Qwen3.6 35B-A3B

Qwen3.6-35B-A3B is a sparse mixture-of-experts model built to balance high-level performance with operational efficiency. By utilizing 35 billion total parameters while activating only 3 billion per token, the architecture achieves a lightweight footprint that allows it to rival much larger dense models in complex environments. Its design centers on agentic coding, featuring specialized capabilities for repository-level reasoning, frontend workflows, and multi-step tool calling. This makes it a versatile choice for developers who require deep analytical power without the resource demands typically associated with dense, large-scale systems.

The model benefits from a post-training lineage that emphasizes stability and real-world utility, incorporating direct community feedback to refine its responsiveness. It features a sophisticated hidden layout that integrates gated DeltaNet and gated attention mechanisms, supporting both multimodal perception and a new thinking preservation option that retains reasoning context across historical messages. These advancements streamline iterative development and improve precision in coding tasks. With broad compatibility across standard frameworks and hardware, the model is positioned as a robust tool for building next-generation agentic workflows that demand both speed and logical depth.

Alibabaqwen3.6-35b-a3bqwen

Quick Info

Powered by
Provider
Alibaba
Model key
qwen3.6-35b-a3b
Release date
Apr 17, 2026
Last updated
Apr 17, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.248
Output token cost
$1.485

Limits

Output tokens
65,536 tokens
Context window
262,144 tokens

Transparent token rates

Compare Qwen3.6 35B-A3B pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Qwen3.6 35B-A3B

Alibaba

Coverage

The news blog specialized in Japanese culture, odd news, gadgets and all other funny stuffs. Updated everyday.

Alibaba

Coverage

On April 2, 2026, Alibaba's Qwen team released Qwen3.6-35B-A3B as an open-weight, Apache 2.0 licensed variant of the Qwen3.6 generation on Hugging Face, paired with the closed Qwen3.6-Plus API model. The 35-billion-parameter Mixture-of-Experts model activates only 3B parameters per token and was framed with the tagline Qwen3.6-35B-A3B uses a sparse MoE with 256 experts (8 routed plus 1 shared activating per token), a 40-layer stack of three Gated DeltaNet linear-attention layers followed by one Gated Attention layer, a hidden dimension of 2048, an expert intermediate dimension of 512, and Multi-Token Prediction training for faster sp

Videos about Qwen3.6 35B-A3B

More models around Qwen3.6 35B-A3B