Currently listed through these providers:
Model details
Qwen3.6 35B-A3B
Built on the Qwen lineage as a sparse mixture-of-experts design, this 35-billion-total-parameter model activates only about 3 billion parameters per token, pairing a large knowledge capacity with the inference efficiency typical of MoE routing. The release framing positions it as a successor to the prior 35B-A3B generation, with a stated emphasis on agentic coding workflows where it is reported to compete with substantially larger dense models. It continues to support both thinking and non-thinking operating modes, and is described as versatile across multimodal perception, reasoning, and structured generation tasks.
In practical use, the model fits developers and teams who want a single open-weight checkpoint for long-context work such as document analysis, code generation, and tool-driven agents, rather than juggling several specialized models. Its sparse activation profile makes it well suited for sustained agentic loops and structured-output pipelines where consistent latency and predictable tool calling matter. For organizations already standardizing on an OpenAI-compatible client surface, the model can be wired in with minimal integration effort, keeping the focus on task design rather than plumbing.
Quick Info
Powered by- Provider
- Zenifra
- Model key
- alibaba/qwen3.6-35b-a3b
- Release date
- Apr 17, 2026
- Last updated
- Apr 17, 2026
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.19
- Output token cost
- $0.48
Limits
- Output tokens
- 65,536 tokens
- Context window
- 262,144 tokens