Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Kilo Gateway logo

Model details

Qwen: Qwen3 VL 30B A3B Instruct

Qwen3 VL 30B A3B Instruct is a multimodal large language model built around a Mixture-of-Experts architecture, with roughly 31.1 billion total parameters and about 3 billion active per token during inference. This MoE design allows the model to scale efficiently across different deployment environments, from edge devices to cloud infrastructure, while maintaining strong performance in both text generation and visual understanding. The architecture includes innovations like Interleaved-MRoPE and DeepStack fusion that enable unified processing of text, images, and video within a single pipeline. The Instruct variant is specifically tuned for instruction-following across general multimodal tasks, excelling in spatial reasoning, OCR across 32 languages, GUI automation, visual coding from sketches to debugged interfaces, and long-context comprehension that handles hours of video or entire documents with full recall.

As an open-weight model released under the Apache 2.0 license, Qwen3 VL 30B A3B is freely accessible and designed for customization. The Instruct variant is optimized through instruction-tuning to follow multi-step directives, handle multi-image inputs, and maintain coherent multi-turn conversations across visual contexts. The underlying MoE architecture supports fine-tuning techniques like LoRA, making it practical to adapt the model for specialized workflows without requiring full retraining. Text performance matches flagship Qwen3 language models, giving it an edge in document AI, spatial tasks, and agent research where visual reasoning and language generation must work together seamlessly. The combination of open accessibility, instruction-tuned behavior, and a long context window makes this model particularly suitable for teams building custom multimodal agents or automating complex visual workflows.

Kilo Gatewayqwen/qwen3-vl-30b-a3b-instructqwen

Quick Info

Powered by
Provider
Kilo Gateway
Model key
qwen/qwen3-vl-30b-a3b-instruct
Release date
Oct 6, 2025
Last updated
Oct 6, 2025
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.13
Output token cost
$0.52

Limits

Output tokens
16,384 tokens
Context window
262,144 tokens

Transparent token rates

Compare Qwen: Qwen3 VL 30B A3B Instruct pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Qwen: Qwen3 VL 30B A3B Instruct

No articles yet. Fetch the latest news to show it here.

Videos about Qwen: Qwen3 VL 30B A3B Instruct

More models around Qwen: Qwen3 VL 30B A3B Instruct