Currently listed through these providers:
Model details
Qwen: Qwen3 VL 30B A3B Instruct
Qwen3 VL 30B A3B Instruct is a multimodal large language model built around a Mixture-of-Experts architecture, with roughly 31.1 billion total parameters and about 3 billion active per token during inference. This MoE design allows the model to scale efficiently across different deployment environments, from edge devices to cloud infrastructure, while maintaining strong performance in both text generation and visual understanding. The architecture includes innovations like Interleaved-MRoPE and DeepStack fusion that enable unified processing of text, images, and video within a single pipeline. The Instruct variant is specifically tuned for instruction-following across general multimodal tasks, excelling in spatial reasoning, OCR across 32 languages, GUI automation, visual coding from sketches to debugged interfaces, and long-context comprehension that handles hours of video or entire documents with full recall.
As an open-weight model released under the Apache 2.0 license, Qwen3 VL 30B A3B is freely accessible and designed for customization. The Instruct variant is optimized through instruction-tuning to follow multi-step directives, handle multi-image inputs, and maintain coherent multi-turn conversations across visual contexts. The underlying MoE architecture supports fine-tuning techniques like LoRA, making it practical to adapt the model for specialized workflows without requiring full retraining. Text performance matches flagship Qwen3 language models, giving it an edge in document AI, spatial tasks, and agent research where visual reasoning and language generation must work together seamlessly. The combination of open accessibility, instruction-tuned behavior, and a long context window makes this model particularly suitable for teams building custom multimodal agents or automating complex visual workflows.
Quick Info
Powered by- Provider
- Kilo Gateway
- Model key
- qwen/qwen3-vl-30b-a3b-instruct
- Release date
- Oct 6, 2025
- Last updated
- Oct 6, 2025
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.13
- Output token cost
- $0.52
Limits
- Output tokens
- 16,384 tokens
- Context window
- 262,144 tokens
Transparent token rates
Compare Qwen: Qwen3 VL 30B A3B Instruct pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about Qwen: Qwen3 VL 30B A3B Instruct
No articles yet. Fetch the latest news to show it here.