Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
NovitaAI logo

Model details

DeepSeek R1 Distill Qwen 32B

DeepSeek R1 Distill Qwen 32B is a distilled large language model that takes the Qwen 2.5 32B base and fine-tunes it on reasoning traces generated by DeepSeek R1, pairing a capable open base with the larger model's chain-of-thought style training signal. The result is a text-to-text system positioned for reasoning-heavy workloads such as multi-step analysis, instruction following, and code-related assistance where step-by-step thinking improves quality. Because the model is presented as a distillate rather than a from-scratch training run, its behavior leans toward the patterns and conventions of the Qwen 2.5 family while inheriting R1-style reasoning habits, making it useful as a practical middle ground between smaller instruction-tuned models and frontier reasoning systems.

The distilled model is offered for local and self-hosted inference alongside hosted access, with community GGUF quantizations available for runtimes such as llama.cpp, including configurations designed to fit within the memory budget of a single high-end consumer GPU. It is marketed as outperforming OpenAI's o1-mini on relevant reasoning evaluations, reflecting the strength of the R1-derived supervision signal rather than a larger parameter count. In practice this version fits teams that want stronger reasoning than a typical 7B to 14B class model without moving to the largest frontier systems, and it pairs naturally with tool-calling and temperature-controlled decoding for agent-style or code-generation pipelines.

NovitaAIdeepseek/deepseek-r1-distill-qwen-32bdeepseek-thinking

Quick Info

Powered by
Provider
NovitaAI
Model key
deepseek/deepseek-r1-distill-qwen-32b
Release date
Jan 20, 2025
Last updated
Jan 20, 2025
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.30
Output token cost
$0.30

Limits

Output tokens
32,000 tokens
Context window
64,000 tokens

Transparent token rates

Compare DeepSeek R1 Distill Qwen 32B pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about DeepSeek R1 Distill Qwen 32B

No articles yet. Fetch the latest news to show it here.

Videos about DeepSeek R1 Distill Qwen 32B

More models around DeepSeek R1 Distill Qwen 32B