Sulat.com
AI models
Get 10-25% off
Get 10-25% off from Qwen
Alibaba logo

Model details

Qwen3 14B

Qwen3-14B is a dense 14.8 billion parameter causal language model that marks a significant step forward in the Qwen series. What sets it apart is its ability to seamlessly switch between a thinking mode for complex logical reasoning, mathematics, and coding, and a non-thinking mode for efficient, general-purpose dialogue—all within a single model. The architecture features 40 layers with grouped query attention (40 heads for queries, 8 for key-value pairs), and supports a native context length of 32,768 tokens. This flexibility allows developers to handle varied workloads without deploying separate models, whether they need deep analytical problem-solving or quick conversational responses.

Built through both pretraining and post-training stages, Qwen3-14B delivers reasoning capabilities that surpass its predecessors, including earlier QwQ thinking models and Qwen2.5 instruct variants, across mathematics, code generation, and commonsense logical reasoning. The model excels in agent-based tasks, enabling precise integration with external tools in both thinking and unthinking modes—achieving leading performance among open-source models in complex agent workflows. It demonstrates strong human preference alignment, performing well in creative writing, role-playing, multi-turn dialogues, and instruction following. With support for over 100 languages and dialects, Qwen3-14B is available as open weights under the Apache 2.0 license, with quantized GGUF versions offering more accessible deployment options for resource-constrained environments.

Alibabaqwen3-14bqwen

Quick Info

Powered by
Provider
Alibaba
Model key
qwen3-14b
Release date
Apr 1, 2025
Last updated
Apr 1, 2025
Knowledge cutoff
2025-04
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.35
Output token cost
$1.40

Limits

Output tokens
8,192 tokens
Context window
131,072 tokens

Transparent token rates

Compare Qwen3 14B pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Qwen3 14B

No articles yet. Fetch the latest news to show it here.

Videos about Qwen3 14B

More models around Qwen3 14B