Sulat.com
AI models
Get 10-25% off
Get 10-25% off from Qwen
Alibaba logo

Model details

Qwen3-Coder 480B-A35B Instruct

Qwen3-Coder-480B-A35B-Instruct is a specialized causal language model built on a Mixture-of-Experts architecture, designed to serve as a highly capable assistant for software development. By utilizing 160 experts with 8 activated during each inference, the model achieves a balance between massive scale and computational efficiency. It is engineered specifically for agentic coding workflows, featuring a dedicated function call format that allows it to interact autonomously with developer environments and tools. This design intent focuses on enabling the model to handle multi-step coding tasks and generate functional applications, positioning it as a competitive alternative to proprietary models in complex programming scenarios.

The model underwent comprehensive pre-training and post-training stages to refine its reasoning and coding capabilities, resulting in state-of-the-art performance on benchmarks like SWE-bench-Verified. Its architecture, which includes 62 layers and 96 attention heads, supports a native context window of 256,000 tokens that can be extended up to 1 million tokens using Yarn, making it particularly effective for understanding large codebases. By focusing on agentic browser-use and foundational coding tasks, the model provides a robust foundation for developers seeking to automate intricate engineering workflows. Its ability to perform at a high level while maintaining an open-weight structure makes it a significant advancement for collaborative and autonomous software engineering.

Alibabaqwen3-coder-480b-a35b-instructqwen

Quick Info

Powered by
Provider
Alibaba
Model key
qwen3-coder-480b-a35b-instruct
Release date
Apr 1, 2025
Last updated
Apr 1, 2025
Knowledge cutoff
2025-04
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$1.50
Output token cost
$7.50

Limits

Output tokens
65,536 tokens
Context window
262,144 tokens

Transparent token rates

Compare Qwen3-Coder 480B-A35B Instruct pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Qwen3-Coder 480B-A35B Instruct

No articles yet. Fetch the latest news to show it here.

Videos about Qwen3-Coder 480B-A35B Instruct

More models around Qwen3-Coder 480B-A35B Instruct