Sulat.com
AI models
Get 10-25% off
Get 10-25% off from Qwen
Alibaba logo

Model details

Qwen3 Coder Flash

Qwen3 Coder Flash is built on a sparse mixture-of-experts architecture with 30.5B total parameters and 3.3B active parameters at any one time, a design that allows it to run smoothly on a 64GB Mac and even on a 32GB Mac when quantized. This non-thinking model is purpose-built for coding tasks, combining a compact active-parameter footprint with full coding proficiency across many programming languages. It operates as a lightweight agent, specializing in autonomous programming through environment interaction, tool use, and agentic coding workflows, making it well-suited for real-time IDE integration and high-volume coding assistance without the overhead of extended reasoning cycles.

The model represents a distillation of Alibaba's larger Qwen3 Coder Plus family, retaining strong coding performance while trimming down to a portable scale. It matches or surpasses leading open-source alternatives on agentic coding benchmarks and tool-use tasks, sitting just behind the flagship 480B flagship version as well as top closed models like Claude Sonnet-4 and GPT-4.1. Its YaRN-based context extension allows it to natively process 256K tokens with room to scale further, helping developers work with entire project libraries without the fragmentation that typically comes with narrower context windows. The combination of agent capability, multi-platform support, and a specially designed function call format positions it as a practical everyday coding partner for developers who need strong performance in a lightweight, deployable package.

Alibabaqwen3-coder-flashqwen

Quick Info

Powered by
Provider
Alibaba
Model key
qwen3-coder-flash
Release date
Jul 28, 2025
Last updated
Jul 28, 2025
Knowledge cutoff
2025-04
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.30
Output token cost
$1.50

Limits

Output tokens
65,536 tokens
Context window
1,000,000 tokens

Transparent token rates

Compare Qwen3 Coder Flash pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Qwen3 Coder Flash

Alibaba

CoverageBenchmark

OpenCode's Alibaba model directory lists Qwen3 Coder Flash with metadata indicating a 1M-token context window, 66K maximum output, and a release date of July 28, 2025. This corroborates the subject's identity and launch timing from a second independent aggregator, complementing the n8n benchmark entry. The page aggrega Notably, the recent usage-share column for Qwen3 Coder Flash is blank, suggesting no consuming token volume in the OpenCode dataset, while sister models such as Qwen3.7 Plus (65%), Qwen3.7 Max (9.87%), and Qwen3.8 Max (8.12%) show active share. The aggregated 2.9T tokens processed across Alibaba models and the 1% recen

Alibaba

CoverageBenchmark

The n8n AI Benchmark page provides a dedicated technical profile for Qwen3 Coder Flash, describing it as Alibaba's fast, cost-efficient proprietary coding agent model based on Qwen3 Coder Plus. It is specialized for autonomous programming via tool calling and environment interaction, combining coding proficiency with g The page reports benchmark sub-scores positioning Qwen3 Coder Flash 8th overall with an overall score of 77, with standout marks for cost-efficiency (97, 4th) and speed (95, 5th). Other recorded scores include Logic 77, Hallucination 80, Structured Output 67, Tool Use 32, Classification 39, and Scoring 49. The listed m

Videos about Qwen3 Coder Flash

More models around Qwen3 Coder Flash