Currently listed through these providers:
Model details
Qwen3 235B-A22B Instruct 2507
Qwen3 235B A22B Instruct 2507 is built on a Mixture-of-Experts architecture that routes each token through 8 of 128 expert layers, activating only 22 billion of its 235 billion total parameters during inference. This sparsity means serving costs stay proportional to active compute while the full parameter space retains the breadth of knowledge needed to compete with larger proprietary models. The model supports 119 languages and dialects and includes a hybrid reasoning system offering both extended "thinking mode" for step-by-step problem solving and a non-thinking mode that delivers immediate responses. Tool calling and Model Context Protocol support make it suitable for multi-step workflows where the model orchestrates external tools, with Alibaba recommending the Qwen-Agent framework for complex agentic pipelines.
This instruction-tuned variant builds on the base Qwen3 235B architecture with post-training refinements that deliver substantial improvements in instruction following, logical reasoning, text comprehension, mathematics, science, coding, and tool usage compared to its predecessor. The model shows markedly better alignment with user preferences in subjective and open-ended tasks, enabling more helpful responses and higher-quality text generation. Benchmark results demonstrate competitive performance against leading models, including strong scores on mathematics reasoning benchmarks, coding assessments, and alignment evaluations. Its native 256K context window with extension to 1M tokens makes it effective for long-document understanding and complex multi-turn conversations, while the ability to configure a thinking budget per request allows dynamic tuning of the cost-quality tradeoff for different use cases.
Quick Info
Powered by- Provider
- Cortecs
- Model key
- qwen3-235b-a22b-instruct-2507
- Release date
- Jul 21, 2025
- Last updated
- Jul 21, 2025
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.069
- Output token cost
- $0.455
Limits
- Output tokens
- 131,000 tokens
- Context window
- 262,000 tokens