Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
OpenRouter logo

Model details

DeepSeek-R1

DeepSeek-R1 was developed as part of DeepSeek's first-generation reasoning line, alongside the sibling variant DeepSeek-R1-Zero. R1-Zero was trained using large-scale reinforcement learning directly on the base model without any preliminary supervised fine-tuning, which let reasoning behaviors emerge naturally but produced outputs with readability and language-mixing issues. R1 addresses those limitations through a multi-stage pipeline that starts with a cold-start phase, continues with reasoning-oriented reinforcement learning, then uses rejection sampling and supervised fine-tuning before a final reinforcement stage aimed at broader scenarios. The release also open-sourced six dense distilled models ranging from 1.5B to 70B parameters, built on Qwen and Llama bases, so smaller deployments can inherit part of R1's reasoning ability.

The hosted variant served through the OpenRouter router is a 671B-parameter mixture-of-experts architecture that activates around 37B parameters per inference pass, making it a heavyweight reasoning system rather than a lightweight chat model. On math, code, and general reasoning benchmarks reported by the DeepSeek team, R1 reaches performance comparable to OpenAI's o1-1217, and the May 28th revision tightens that comparison further while exposing its full reasoning traces for inspection. For practitioners it fits well in workflows that need chain-of-thought inspection, tool integration, and structured outputs on long contexts, and its open weights combined with permissive licensing make it attractive for teams that want to study or self-host a frontier-class reasoning model without paying closed-API premiums.

OpenRouterdeepseek/deepseek-r1deepseek-thinking

Quick Info

Powered by
Provider
OpenRouter
Model key
deepseek/deepseek-r1
Release date
Jan 20, 2025
Last updated
May 29, 2025
Knowledge cutoff
2024-07
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.70
Output token cost
$2.50

Limits

Output tokens
16,000 tokens
Context window
64,000 tokens

Transparent token rates

Compare DeepSeek-R1 pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about DeepSeek-R1

Vercel AI Gateway

CoverageBenchmark

A peer-reviewed meta-analysis published in the Journal of Big Data (Volume 13, article 26, 2026; published 19 December 2025) directly benchmarks DeepSeek-R1 against GPT-4 Turbo, Gemini Ultra, Qwen, and LLaMA 3.1 using standardized tasks including MMLU, HumanEval, FLORES-200, and TyDiQA. The hybrid meta-analysis aggrega Reported results for DeepSeek-R1 include 80.2 ± 1.5% on HumanEval and 78.5 ± 1.8% on MMLU, compared with ChatGPT-4 Turbo's 86.5 ± 1.9% on HumanEval, with the gap falling within observed heterogeneity (I² = 14.6%). The study concludes R1 demonstrates strong coding and multilingual efficiency, trails GPT-4 Turbo in reaso

OpenRouter

Official sourceBenchmark

DeepSeek R1 on OpenRouter is priced at $0.70 per 1M input tokens and $2.50 per 1M output tokens, with a 164K context window, and was released on January 20, 2025 with a July 2024 knowledge cutoff. The model is a 671B-parameter mixture-of-experts with 37B active parameters per inference pass, MIT licensed for distillati OpenRouter currently routes DeepSeek R1 traffic primarily to NovitaAI, which reports 100% uptime, 0.94s P50 latency, and 18 tok/s throughput at standard routing, with weighted-average effective input pricing of $0.6999/M tokens. Azure is listed as a secondary provider but shows 0% uptime in the current window, and Novi

OpenRouter

Coverage

The official DeepSeek API changelog (api-docs.deepseek.com/updates/) documents that as of 2026-04-24 the DeepSeek API supports V4-Pro and V4-Flash via both the OpenAI ChatCompletions and Anthropic interfaces, with the base URL unchanged and model parameters set to deepseek-v4-pro or deepseek-v4-flash. The legacy aliase During the transition window the legacy deepseek-chat and deepseek-reasoner names point to the non-thinking and thinking modes of deepseek-v4-flash respectively, meaning the OpenRouter-hosted DeepSeek-R1 (model key deepseek/deepseek-r1) in the DeepSeek thinking-model family is affected by this same deprecation. Develop

OpenRouter

Official sourceComparison

Compare R1 Distill Qwen 32B from DeepSeek and R1 from DeepSeek on key metrics including benchmarks, price, context length, and other model features.

OpenRouter

Official sourceBenchmark

May 28th update to the [original DeepSeek R1](/deepseek/deepseek-r1) Performance on par with [OpenAI o1](/openai/o1), but open-sourced and with fully open reasoning tokens. $0.50 per million input tokens, $2.15 per million output tokens. 163,840 token context window, maximum output of 32,768 tokens. Higher uptime with

OpenRouter

Coverage

Official DeepSeek reasoning model first released on January 20, 2025 and updated on May 28, 2025.

OpenRouter

Coverage

Well that didn't take long, available from 7 providers through openrouter. https://openrouter.ai/deepseek/deepseek-r1-0528/providers. May 28th update to the ...

Videos about DeepSeek-R1

More models around DeepSeek-R1