Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Hugging Face logo

Model details

DeepSeek-R1-0528

DeepSeek R1-0528 is a May 2025 refresh of the original R1 reasoning model, refining an already strong open-weight foundation with better reasoning depth, inference efficiency, and benchmark gains. It sits in the deepseek-thinking family and continues the lineage that positions R1 as a flagship reasoning system, with this update specifically targeting improvements in logic, mathematics, and programming tasks. The model is offered under a fully open-source license with openly exposed reasoning tokens, making it unusual among top-tier reasoning systems where the chain-of-thought is typically hidden.

At the architecture level, R1-0528 is a Mixture-of-Experts model with 671 billion total parameters but only 37 billion active during any single inference pass, which keeps per-query compute manageable while preserving broad capability. Independent benchmarking coverage describes it as the most capable open-weight model of its release window, approaching the performance tier of leading closed frontier systems. Practical fit is strongest for teams that need deep, transparent reasoning on long-context problems and want full control over the weights, since the open distribution allows local deployment, distillation into smaller variants, and integration into custom pipelines without depending on a single hosted provider.

Hugging Facedeepseek-ai/DeepSeek-R1-0528deepseek-thinking

Quick Info

Powered by
Provider
Hugging Face
Model key
deepseek-ai/DeepSeek-R1-0528
Release date
May 28, 2025
Last updated
May 28, 2025
Knowledge cutoff
2025-05
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$3.00
Output token cost
$5.00

Limits

Output tokens
163,840 tokens
Context window
163,840 tokens

Transparent token rates

Compare DeepSeek-R1-0528 pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about DeepSeek-R1-0528

Hugging Face

CoverageAnalysis

Runpod's "Minor Upgrade" deep dive, published August 25, 2026, characterizes DeepSeek-R1-0528 as a substantial stealth release that continues to use a Mixture-of-Experts architecture scaled up to a very large parameter count while preserving a 128K context window (extendable via RoPE scaling). The piece contextualizes The deep dive documents the headline mathematical-reasoning gains on AIME 2025 (accuracy rising from 70% in the prior version to 87.5% in R1-0528, putting it near OpenAI o3 at 88.9% and ahead of Gemini 2.5 Pro at 83.0%). The article attributes the gain to "enhanced thinking depth," noting average reasoning tokens per A

Hugging Face

CoverageRelease Notes

DeepSeek-R1-0528 is an updated version of DeepSeek-R1 with improved reasoning, inference, and performance via optimizations and enhanced computational...

Hugging Face

Coverage

The Chinese start-up DeepSeek has updated its R1 model, improving its performance in reasoning, logic, mathematics, and programming. This update, which reduces...

Hugging Face

CoverageBenchmark

DeepSeek’s New R1–0528: Performance Analysis and Benchmark Comparisons TL;DR: The newest DeepSeek R1 model is the most powerful among open-weight models, approaching performance of the leading …

Videos about DeepSeek-R1-0528

More models around DeepSeek-R1-0528