Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Deep Infra logo

Model details

DeepSeek-R1-0528

DeepSeek-R1-0528 is a late-May 2025 refresh of the original DeepSeek-R1, built to push the same reasoning-first lineage further rather than reinvent it. DeepSeek framed the release around measurable reasoning gains, and The Verge notes that the update significantly improved the model's depth of reasoning and inference capabilities while reducing hallucination rates compared with the prior version. Fireworks AI's listing echoes that positioning, describing the checkpoint as approaching the overall performance of leading frontier systems. The model also broadens its practical utility with explicit support for structured JSON output and function calling, making it more suitable for tool-augmented agents and pipelines that need dependable structured responses.

In real-world use, DeepSeek-R1-0528 fits scenarios that demand careful multi-step thinking, such as analytical Q&A, code generation and debugging, and workflow automations that hand off between the model and external tools. The combination of stronger reasoning, cleaner structured output, and tighter function-calling behavior makes it a good match for developers building AI assistants or backend services rather than lightweight chat use. Open weights are published on Hugging Face for teams that want to self-host or inspect the model, and broader availability through channels like GitHub Models has made it easy to evaluate alongside other reasoning systems. For practitioners deciding between R1 versions, the 0528 checkpoint is essentially the same accessible but more dependable reasoning model, with fewer factual slips and more reliable tool integration.

Deep Infradeepseek-ai/DeepSeek-R1-0528

Quick Info

Powered by
Provider
Deep Infra
Model key
deepseek-ai/DeepSeek-R1-0528
Release date
May 28, 2025
Last updated
May 28, 2025
Knowledge cutoff
2024-07
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.50
Output token cost
$2.15

Limits

Output tokens
64,000 tokens
Context window
163,840 tokens

Latest news about DeepSeek-R1-0528

Hugging Face

CoverageAnalysis

Runpod's "Minor Upgrade" deep dive, published August 25, 2026, characterizes DeepSeek-R1-0528 as a substantial stealth release that continues to use a Mixture-of-Experts architecture scaled up to a very large parameter count while preserving a 128K context window (extendable via RoPE scaling). The piece contextualizes The deep dive documents the headline mathematical-reasoning gains on AIME 2025 (accuracy rising from 70% in the prior version to 87.5% in R1-0528, putting it near OpenAI o3 at 88.9% and ahead of Gemini 2.5 Pro at 83.0%). The article attributes the gain to "enhanced thinking depth," noting average reasoning tokens per A

Deep Infra

CoverageRelease Notes

DeepSeek-R1-0528 is an updated version of DeepSeek-R1 with improved reasoning, inference, and performance via optimizations and enhanced computational...

Deep Infra

Coverage

This development reportedly intensifies the competition within the AI sector, particularly with US-based firms such as OpenAI and Google.

Deep Infra

Coverage

The Chinese AI model that shook up the industry as a more cost-efficient alternative to the ones from OpenAI, Google, and Meta, now has a new update dubbed DeepSeek-R1-0528. DeepSeek says its latest model has a reduced “hallucination” rate, and that it “has significantly improved its depth of reasoning and inference ca

Deep Infra

CoverageBenchmark

Analysis of DeepSeek's DeepSeek R1 0528 (May '25) and comparison to other AI models across key metrics including quality, price, performance (tokens per second & time to first token), context window & more.

Videos about DeepSeek-R1-0528