Currently listed through these providers:
Model details
Qwen: Qwen3 30B A3B Instruct 2507
Qwen3 30B A3B Instruct 2507 is a Mixture-of-Experts causal language model built around a sparse architecture that activates only a fraction of its total parameters during inference. With 128 routing experts and 8 activated per forward pass, the model achieves efficient computation while maintaining broad capability coverage across its 30.5 billion total parameters. The design reflects a deliberate tradeoff: rather than activating all weights for every token, the model dynamically routes each computation through specialized expert sub-networks, allowing it to specialize across diverse knowledge domains without proportional compute cost. Operating exclusively in non-thinking mode, this variant skips intermediate reasoning traces and delivers direct responses, making it particularly suited for applications where latency and concise output matter more than showing step-by-step deliberation.
The 2507 update represents a significant post-training iteration focused on aligning the model more closely with practical user needs. Improvements center on instruction following, mathematical reasoning, coding tasks, and tool usage capabilities, alongside stronger multilingual coverage for long-tail knowledge across dozens of languages. The model also demonstrates meaningfully better performance on subjective and open-ended tasks where user preference alignment is difficult to measure mechanistically. Third-party security evaluators have conducted red teaming analysis across diverse attack vectors, providing external scrutiny of the model's safety profile as adoption grows. The combination of open weights availability and FP8 quantization support makes this model accessible for researchers and developers who want to inspect, fine-tune, or deploy it with reasonable hardware constraints.
Quick Info
Powered by- Provider
- Kilo Gateway
- Model key
- qwen/qwen3-30b-a3b-instruct-2507
- Release date
- Jul 29, 2025
- Last updated
- Jul 29, 2025
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.13
- Output token cost
- $0.52
Limits
- Output tokens
- 32,000 tokens
- Context window
- 128,000 tokens
Transparent token rates
Compare Qwen: Qwen3 30B A3B Instruct 2507 pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about Qwen: Qwen3 30B A3B Instruct 2507
No articles yet. Fetch the latest news to show it here.