Currently listed through these providers:
Model details
Qwen3-Coder 480B-A35B Instruct
Qwen3-Coder-480B-A35B-Instruct is a specialized causal language model built on a Mixture-of-Experts architecture, designed specifically to excel at agentic coding and complex software development tasks. By utilizing 160 total experts with 8 active experts per inference, the model achieves a balance between high-level reasoning and computational efficiency. It is engineered to support autonomous interactions within developer environments, featuring a custom function-calling format that allows it to interface effectively with platforms like CLINE and Qwen Code to manage multi-step workflows.
The model underwent extensive pre-training and post-training phases to refine its performance, resulting in capabilities that compete with leading proprietary models on foundational coding benchmarks such as SWE-bench-Verified. With native support for a 256,000-token context window that can be extended to 1 million tokens using Yarn, the model is optimized for deep repository-scale understanding. This combination of massive parameter scale and architectural optimization makes it a robust choice for developers seeking an open-source solution for building functional applications and navigating intricate, large-scale codebases.
Quick Info
Powered by- Provider
- DevPass (LLM Gateway)
- Model key
- qwen3-coder-480b-a35b-instruct
- Release date
- Apr 1, 2025
- Last updated
- Apr 1, 2025
- Knowledge cutoff
- 2025-04
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.38
- Output token cost
- $1.55
Limits
- Output tokens
- 65,536 tokens
- Context window
- 262,144 tokens
Transparent token rates
Compare Qwen3-Coder 480B-A35B Instruct pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about Qwen3-Coder 480B-A35B Instruct
No articles yet. Fetch the latest news to show it here.