Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Kilo Gateway logo

Model details

Qwen: Qwen3 30B A3B Instruct 2507

Qwen3 30B A3B Instruct 2507 is a Mixture-of-Experts causal language model built around a sparse architecture that activates only a fraction of its total parameters during inference. With 128 routing experts and 8 activated per forward pass, the model achieves efficient computation while maintaining broad capability coverage across its 30.5 billion total parameters. The design reflects a deliberate tradeoff: rather than activating all weights for every token, the model dynamically routes each computation through specialized expert sub-networks, allowing it to specialize across diverse knowledge domains without proportional compute cost. Operating exclusively in non-thinking mode, this variant skips intermediate reasoning traces and delivers direct responses, making it particularly suited for applications where latency and concise output matter more than showing step-by-step deliberation.

The 2507 update represents a significant post-training iteration focused on aligning the model more closely with practical user needs. Improvements center on instruction following, mathematical reasoning, coding tasks, and tool usage capabilities, alongside stronger multilingual coverage for long-tail knowledge across dozens of languages. The model also demonstrates meaningfully better performance on subjective and open-ended tasks where user preference alignment is difficult to measure mechanistically. Third-party security evaluators have conducted red teaming analysis across diverse attack vectors, providing external scrutiny of the model's safety profile as adoption grows. The combination of open weights availability and FP8 quantization support makes this model accessible for researchers and developers who want to inspect, fine-tune, or deploy it with reasonable hardware constraints.

Kilo Gatewayqwen/qwen3-30b-a3b-instruct-2507qwen

Quick Info

Powered by
Provider
Kilo Gateway
Model key
qwen/qwen3-30b-a3b-instruct-2507
Release date
Jul 29, 2025
Last updated
Jul 29, 2025
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.13
Output token cost
$0.52

Limits

Output tokens
32,000 tokens
Context window
128,000 tokens

Transparent token rates

Compare Qwen: Qwen3 30B A3B Instruct 2507 pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Qwen: Qwen3 30B A3B Instruct 2507

No articles yet. Fetch the latest news to show it here.

Videos about Qwen: Qwen3 30B A3B Instruct 2507

More models around Qwen: Qwen3 30B A3B Instruct 2507