Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
SiliconFlow logo

Model details

Qwen/Qwen3-14B

Qwen3-14B is the 14-billion-parameter dense member of the Qwen3 family, representing a major step forward from Qwen2.5. Built on a foundation of 36 trillion tokens of pre-training data across 119 languages—tripling the language coverage of its predecessor—the model brings broad multilingual capability alongside deep expertise in STEM, coding, and logical reasoning. Its architecture incorporates a series of training refinements, including qk layernorm for improved stability. The model supports both thinking and non-thinking modes, allowing users to toggle reasoning behavior as needed, and excels at creative writing, role-playing, multi-turn dialogues, and instruction following. Advanced agent capabilities extend its practical utility into task-oriented workflows.

Qwen3-14B undergoes a structured three-stage pre-training process: broad language modeling and knowledge acquisition first, followed by targeted reasoning skill development in STEM and coding domains, then long-context training extending to 32k token sequences. Scaling law guided hyperparameter tuning across the pipeline separately optimizes critical settings for dense models like this one, improving overall training dynamics and final performance. With support for over 100 languages and dialects, plus tool-calling readiness built into the training, the model is positioned for versatile deployment across both conversational and agentic use cases. This combination of multilingual breadth, reasoning flexibility, and agent-ready design makes the 14B variant a practical mid-scale choice for teams needing strong general capabilities without the overhead of much larger models.

SiliconFlowQwen/Qwen3-14Bqwen

Quick Info

Powered by
Provider
SiliconFlow
Model key
Qwen/Qwen3-14B
Release date
Apr 30, 2025
Last updated
Nov 25, 2025
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.07
Output token cost
$0.28

Limits

Output tokens
131,000 tokens
Context window
131,000 tokens

Transparent token rates

Compare Qwen/Qwen3-14B pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Qwen/Qwen3-14B

No articles yet. Fetch the latest news to show it here.

Videos about Qwen/Qwen3-14B

More models around Qwen/Qwen3-14B