Currently listed through these providers:
Model details
Qwen3-14B
Qwen3-14B is a dense causal language model built within the Qwen3 family, Alibaba's latest generation of large language models. With 14.8 billion total parameters distributed across 40 transformer layers and grouped query attention (40 heads for queries, 8 for key-value pairs), the model achieves strong performance while remaining compact enough for practical deployment. The architecture supports a native context length of 32,768 tokens and introduces a defining feature: the ability to switch seamlessly between thinking mode for complex logical reasoning, mathematics, and coding tasks, and non-thinking mode for rapid, efficient general-purpose dialogue. This dual-mode design allows a single model to handle both deep analytical problems and everyday conversational needs without requiring separate specialized systems.
The model was developed through extensive pretraining followed by post-training, positioning it as a successor to the Qwen2.5 instruct series. In benchmark evaluations, Qwen3-14B demonstrates reasoning capabilities that surpass its QwQ predecessor in thinking mode and outperform Qwen2.5 instruct models in non-thinking mode across mathematics, code generation, and commonsense logical reasoning tasks. The model excels in human preference alignment, showing particular strength in creative writing, role-playing, and multi-turn dialogues. Its agent capabilities enable precise integration with external tools in both modes, achieving leading performance among open-source models on complex agent-based tasks. Multilingual support extends across more than 100 languages and dialects, making it suitable for diverse global applications. The dense model weights are available under Apache 2.0 licensing, and quantized GGUF versions offer accessible paths for local deployment.
Quick Info
Powered by- Provider
- Vercel AI Gateway
- Model key
- alibaba/qwen-3-14b
- Release date
- Apr 28, 2025
- Last updated
- Apr 1, 2025
- Knowledge cutoff
- 2025-04
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.12
- Output token cost
- $0.24
Limits
- Output tokens
- 16,384 tokens
- Context window
- 40,960 tokens
Transparent token rates
Compare Qwen3-14B pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about Qwen3-14B
No articles yet. Fetch the latest news to show it here.