Vercel AI Gateway
Alibaba has released Qwen3-Coder-Next, an open-source 80B-parameter coding model that activates just 3B parameters per query, scoring 70.6% on SWE-Bench.
Model details
Qwen3 Coder 480B A35B Instruct is a causal language model built on a Mixture-of-Experts architecture, designed specifically to function as an autonomous coding agent. With 480 billion total parameters and 35 billion active parameters across 160 experts, the model balances high-level reasoning with computational efficiency. It features 62 layers and a specialized attention mechanism, allowing it to handle complex workflows that go beyond simple code generation. The model is engineered for repository-scale tasks, offering native support for 256,000 tokens of context, which can be extended up to 1 million tokens using Yarn, making it highly effective for managing large-scale software projects.
Developed through comprehensive pre-training and post-training stages, the model is optimized for agentic coding, browser-use, and tool-use scenarios. It utilizes a specially designed function call format that allows it to integrate seamlessly with developer tools and environments, including the Qwen Code command-line interface. By achieving performance levels comparable to proprietary models on rigorous benchmarks like SWE-bench-Verified and WebArena, it provides a robust foundation for developers seeking to automate multi-turn reasoning and complex software engineering workflows. Its design focuses on delivering state-of-the-art results in autonomous assistance while maintaining the flexibility required for modern, interactive development.
Transparent token rates
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Vercel AI Gateway
Alibaba has released Qwen3-Coder-Next, an open-source 80B-parameter coding model that activates just 3B parameters per query, scoring 70.6% on SWE-Bench.
Vercel AI Gateway
Alibaba's Qwen3 Coder 480B Instruct model is now available on Cerebras.