Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
NovitaAI logo

Model details

Qwen3 30B A3B

Qwen3 30B A3B is a causal language model built on a mixture-of-experts architecture, featuring 30.5 billion total parameters with 3.3 billion parameters activated during inference. Designed to balance depth and efficiency, the model supports a unique operational design that allows users to switch between a specialized thinking mode for complex tasks like mathematics and coding and a non-thinking mode for standard conversational interactions. This dual-mode capability is supported by a 48-layer structure and grouped-query attention, enabling the model to maintain high performance across diverse scenarios ranging from creative writing to intricate logical problem-solving.

The model benefits from extensive pre-training and post-training stages, resulting in significant advancements in instruction following and agent-based task execution. By integrating seamlessly with external tools, it excels in complex, multi-turn environments where precise reasoning is required. Its training lineage emphasizes human preference alignment, ensuring that the model remains engaging and natural in role-playing and dialogue. With support for over 100 languages and dialects, this model is positioned as a robust solution for developers seeking a balance between high-level reasoning performance and the operational efficiency of a smaller activated parameter count.

NovitaAIqwen/qwen3-30b-a3b-fp8

Quick Info

Powered by
Provider
NovitaAI
Model key
qwen/qwen3-30b-a3b-fp8
Release date
Apr 29, 2025
Last updated
Apr 29, 2025
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.09
Output token cost
$0.45

Limits

Output tokens
20,000 tokens
Context window
40,960 tokens

Latest news about Qwen3 30B A3B

No articles yet. Fetch the latest news to show it here.

Videos about Qwen3 30B A3B