Currently listed through these providers:
Model details
Qwen3.6 Flash
Qwen3.6 Flash belongs to the Qwen3.6 family, a generation that also includes the Qwen3.6-Plus and Qwen3.6-35B-A3B variants. The family line is documented through a Qwen blog post describing Qwen3.6-35B-A3B as a sparse mixture-of-experts model with 35 billion total parameters and 3 billion active parameters, built to deliver agentic coding performance competitive with much larger dense models while remaining efficient at inference. That same family positioning frames Flash as a lighter, lower-latency sibling alongside the higher-capacity Plus model, with the Qwen3.6 generation emphasizing multimodal perception, reasoning, and tool-oriented workflows rather than pure chat capability.
Outside the Qwen blog, Qwen3.6 Flash is surfaced in third-party tooling as a vision-capable model that can be compared side by side with Qwen3.6 Plus on tasks like OCR, image captioning, and open-prompt visual reasoning, suggesting practical utility for developers who want a faster multimodal option without stepping up to the Plus tier. The combination of vision support, coding-oriented family lineage, and a Flash-tier speed profile makes it a reasonable fit for agentic assistants, document and screenshot understanding, and lightweight multimodal pipelines where latency and cost matter more than maximum reasoning depth.
Quick Info
Powered by- Provider
- Alibaba Coding Plan (China)
- Model key
- qwen3.6-flash
- Release date
- Apr 27, 2026
- Last updated
- Apr 27, 2026
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.1875
- Output token cost
- $1.125
Limits
- Output tokens
- 65,536 tokens
- Context window
- 1,000,000 tokens
Latest news about Qwen3.6 Flash
Videos about Qwen3.6 Flash
Recent tweets and retweets from Alibaba Coding Plan (China)
More models around Qwen3.6 Flash
This exact model name is also listed by 20 other providers.