Currently listed through these providers:
Model details
DeepSeek R1 Distill Llama 70B
DeepSeek R1 Distill Llama 70B is built on a decoder-only transformer architecture rooted in the Llama family, distilled from the larger DeepSeek R1 reasoning model to bring advanced chain-of-thought capabilities to a more manageable 70-billion parameter footprint. This variant inherits the ability to decompose complex problems and sustain long reasoning chains—a quality the full DeepSeek R1 series developed through large-scale reinforcement learning—while maintaining the accessibility of an open-weight release under the MIT license. The design targets developers and teams that need strong logical reasoning without deploying a massive frontier model.
The model traces its lineage to DeepSeek's progression from R1-Zero, which demonstrated that pure reinforcement learning could produce impressive reasoning behaviors but suffered from issues like repetition and language mixing. DeepSeek R1 addressed those challenges by introducing cold-start data and refined training recipes before applying RL, creating a more usable foundation for distillation. On hardware like Cerebras' wafer-scale systems, this 70B variant has demonstrated inference throughput exceeding 1,500 tokens per second—roughly 57 times faster than conventional GPU deployments—making it practical for real-time reasoning applications that demand both depth and responsiveness.
Quick Info
Powered by- Provider
- DigitalOcean
- Model key
- deepseek-r1-distill-llama-70b
- Release date
- Jan 30, 2025
- Last updated
- Jan 30, 2025
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.99
- Output token cost
- $0.99
Limits
- Output tokens
- 8,192 tokens
- Context window
- 32,678 tokens
Transparent token rates
Compare DeepSeek R1 Distill Llama 70B pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about DeepSeek R1 Distill Llama 70B
No articles yet. Fetch the latest news to show it here.