Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Together AI logo

Model details

Qwen3 Coder 480B A35B Instruct

Qwen3 Coder 480B A35B Instruct was introduced by the Qwen team as their most agentic code model to date, built around a Mixture-of-Experts architecture with 480 billion total parameters and 35 billion active per forward pass, routing through 8 of 160 experts. The variant is optimized for code generation, function calling, tool use, and long-context reasoning across software repositories, and the team positions it for agentic coding, agentic browser-use, and agentic tool-use workloads where it reportedly matches the strongest closed competitors on open-model benchmarks. It is published as an open-weight release under the Qwen organization, which is what allows the FP8 build to be redistributed and served by third-party hosts.

In practical terms, the model handles very large codebases through a 256K-token native context window that can be extended to roughly one million tokens via extrapolation methods, while the API surface exposed through third-party routers reports a 262K working window and supports both text input and text output with temperature control. It pairs with a dedicated command-line agent, Qwen Code, adapted from Gemini Code with custom prompts and function-calling protocols, making it a natural fit for developers who want to drive repository-scale refactors, multi-step debugging, and tool-augmented coding pipelines. The result is a long-horizon coding assistant that combines open weights, a substantial active-parameter budget, and extended context for agent-style development work.

Together AIQwen/Qwen3-Coder-480B-A35B-Instruct-FP8qwendeprecated

Quick Info

Powered by
Provider
Together AI
Model key
Qwen/Qwen3-Coder-480B-A35B-Instruct-FP8
Release date
Jul 23, 2025
Last updated
Jul 23, 2025
Knowledge cutoff
2025-04
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$2.00
Output token cost
$2.00

Limits

Output tokens
262,144 tokens
Context window
262,144 tokens

Transparent token rates

Compare Qwen3 Coder 480B A35B Instruct pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Qwen3 Coder 480B A35B Instruct

Together AI

Official sourceBenchmark

The official Together AI model page for Qwen3-Coder 480B A35B Instruct confirms specifications: 480B total parameters with 35B activated, 256K context length, FP8 quantization, code/chat modality, and support for function calling and JSON mode. The page categorizes the model under coding agents with primary use cases i A critical status signal on this page states "This model is not available on Together's Serverless API," and no pricing is displayed, which together indicate the model has been removed from or is no longer offered via Together AI's serverless inference catalog. Related models listed alongside include MiniMax M3, GLM-5.

Videos about Qwen3 Coder 480B A35B Instruct

More models around Qwen3 Coder 480B A35B Instruct