Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Vercel AI Gateway logo

Model details

Qwen3 Next 80B A3B Instruct

Qwen3 Next 80B A3B Instruct is a foundation model built to prioritize scaling efficiency through a highly sparse Mixture-of-Experts architecture. By activating only 3 billion parameters per inference step, the model achieves a significant reduction in computational cost while maintaining the capacity of an 80-billion parameter system. Its design incorporates Hybrid Attention, which combines Gated DeltaNet and Gated Attention to manage ultra-long context windows effectively. This architecture is specifically engineered to deliver high-throughput performance, making it a robust choice for tasks requiring deep document analysis, complex multi-turn dialogues, and code generation.

The model benefits from advanced stability optimizations, including zero-centered and weight-decayed layer normalization, which support consistent performance during both pre-training and post-training phases. By utilizing Multi-Token Prediction, the model accelerates inference speeds and enhances overall output quality. It is optimized to provide direct, instruction-following responses without visible chain-of-thought traces, positioning it as a reliable tool for agentic workflows, retrieval-augmented generation, and production environments where deterministic results are required. Its ability to match the performance of much larger systems while maintaining superior inference throughput makes it a versatile solution for enterprise-scale applications.

Vercel AI Gatewayalibaba/qwen3-next-80b-a3b-instructqwen

Quick Info

Powered by
Provider
Vercel AI Gateway
Model key
alibaba/qwen3-next-80b-a3b-instruct
Release date
Sep 1, 2025
Last updated
Sep 1, 2025
Knowledge cutoff
2025-04
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.15
Output token cost
$1.20

Limits

Output tokens
262,114 tokens
Context window
262,114 tokens

Transparent token rates

Compare Qwen3 Next 80B A3B Instruct pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Qwen3 Next 80B A3B Instruct

No articles yet. Fetch the latest news to show it here.

Videos about Qwen3 Next 80B A3B Instruct

More models around Qwen3 Next 80B A3B Instruct