Currently listed through these providers:
Model details
Qwen/Qwen3-14B
Qwen3-14B is a dense, 14.8-billion parameter causal language model built to balance high-level reasoning with efficient, general-purpose interaction. Its design centers on a unique dual-mode architecture that allows users to seamlessly switch between a thinking mode, optimized for complex logical inference, mathematics, and coding, and a non-thinking mode for rapid, natural conversation. This versatility makes the model a robust choice for a wide range of applications, from technical problem-solving and agent-based tool integration to creative writing and immersive role-playing.
Developed through comprehensive pre-training and post-training stages, the model demonstrates significant advancements in instruction-following and human preference alignment. It is engineered to handle complex agent-based tasks and supports over 100 languages and dialects, ensuring strong performance in multilingual translation and instruction. With a native context capacity that scales to 131,072 tokens, the model is well-positioned for long-form analysis and multi-turn workflows, providing a flexible and powerful tool for developers and users requiring both depth in reasoning and breadth in linguistic capability.
Quick Info
Powered by- Provider
- SiliconFlow (China)
- Model key
- Qwen/Qwen3-14B
- Release date
- Apr 30, 2025
- Last updated
- Nov 25, 2025
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.07
- Output token cost
- $0.28
Limits
- Output tokens
- 131,000 tokens
- Context window
- 131,000 tokens
Transparent token rates
Compare Qwen/Qwen3-14B pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about Qwen/Qwen3-14B
No articles yet. Fetch the latest news to show it here.