Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
OpenRouter logo

Model details

Qwen3-Next 80B-A3B Instruct

Qwen3-Next 80B-A3B Instruct is described by third-party listings as an 80-billion-parameter sparse mixture-of-experts language model with roughly 3 billion parameters active per token, paired with a hybrid attention design that combines linear Gated DeltaNet layers with standard gated attention. That combination is presented as a deliberate trade between long-context efficiency and standard transformer recall, making the model well suited to extended-document reasoning, multilingual prompts, and coding sessions that benefit from deep context retention. The instruction-tuned variant focuses on practical assistant behavior, so it is positioned less as a base model and more as a ready-to-use conversational and task-following system for developers who want strong long-context handling without retraining.

Independent leaderboard tracking places Qwen3-Next 80B-A3B Instruct in the middle of the pack overall, with average-tier placements in legal, finance, chat, and healthcare categories, while ranking below the top half of current open models on reasoning, coding, tool calling, vision, writing, and math tasks. In practice this means it reads as a balanced generalist rather than a specialist: a sensible default when long context, open weights, and multilingual coverage matter more than top-tier benchmark scores, and a less compelling pick for workloads that demand the strongest reasoning or coding accuracy. Its sparse MoE design also suggests a favorable cost-to-quality profile at inference time, aligning with its role as an accessible long-context assistant for everyday production use.

OpenRouterqwen/qwen3-next-80b-a3b-instructqwen

Quick Info

Powered by
Provider
OpenRouter
Model key
qwen/qwen3-next-80b-a3b-instruct
Release date
Sep 1, 2025
Last updated
Sep 1, 2025
Knowledge cutoff
2025-04
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.10
Output token cost
$1.10

Limits

Output tokens
235,929 tokens
Context window
262,144 tokens

Transparent token rates

Compare Qwen3-Next 80B-A3B Instruct pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Qwen3-Next 80B-A3B Instruct

OpenRouter

Official sourceBenchmark

OpenRouter's official listing for Qwen3-Next-80B-A3B-Instruct confirms the model was released on September 11, 2025, with a 262K context window, a September 2025 knowledge cutoff, and a default listed price of $0.09 per 1M input tokens and $1.10 per 1M output tokens (DeepInfra). The model is described as an instruction OpenRouter routes the same model weights across five providers with varying latency, throughput, and uptime profiles: DeepInfra ($0.09/$1.10, ~65 tps), Alibaba Cloud International ($0.0975/$0.78, ~59 tps), Parasail ($0.10/$1.10 with $0.07 cache read, ~37 tps), Google Vertex ($0.15/$1.20, ~49 tps), and NovitaAI ($0.15/$

Videos about Qwen3-Next 80B-A3B Instruct

More models around Qwen3-Next 80B-A3B Instruct