Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
NovitaAI logo

Model details

Qwen3 235B A22B

Qwen3 235B A22B is a flagship mixture-of-experts model from the Qwen series, built with 235 billion total parameters but activating only 22 billion during inference. This MoE design allows the model to maintain exceptional capability across complex reasoning, mathematics, and coding tasks while keeping computational costs manageable. A defining characteristic is its ability to seamlessly switch between thinking mode, which enables deep logical analysis for intricate problems, and non-thinking mode for efficient general-purpose dialogue—all within a single model. The architecture spans 94 layers with grouped query attention, supporting a 40960-token context window. It also emphasizes strong agent capabilities, enabling precise integration with external tools in both thinking and non-thinking modes, which makes it well-suited for building autonomous workflows.

The model was developed through extensive pretraining followed by post-training alignment stages that refined its reasoning and instruction-following abilities. Qwen3 235B A22B has demonstrated competitive performance against top-tier models such as DeepSeek-R1, o1, and Gemini-2.5-Pro on benchmarks evaluating coding, math, and general capabilities. It excels in multilingual support across more than 100 languages and dialects, with particular strength in human preference alignment for creative writing, role-playing, and multi-turn conversations. An FP8 quantized variant has been released to improve tool use, coding, and logical reasoning capabilities. Being open-weight, it invites community fine-tuning through methods like LoRA, making it accessible for customization in production environments and specialized applications.

NovitaAIqwen/qwen3-235b-a22b-fp8

Quick Info

Powered by
Provider
NovitaAI
Model key
qwen/qwen3-235b-a22b-fp8
Release date
Apr 29, 2025
Last updated
Apr 29, 2025
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.20
Output token cost
$0.80

Limits

Output tokens
20,000 tokens
Context window
40,960 tokens

Latest news about Qwen3 235B A22B

No articles yet. Fetch the latest news to show it here.

Videos about Qwen3 235B A22B