Skip to content
Papers.

DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning research paper by DeepSeek, 2025

DeepSeek · Jan 22, 2025 · Reasoning · 5,691 citations · 462 upvotes

Read on arXiv

What it shows

Reinforcement learning taught a model long step-by-step reasoning on par with OpenAI o1.

Summarised by hand from the abstract.

More from DeepSeek

All 2
Topic
About this paper
Authors
DeepSeek-AI, Daya Guo, Dejian Yang and 197 more
arXiv
2501.12948 · PDF
Venue
Nature
Citations
5,691, 921 influential · Semantic Scholar
Upvotes
462 · Hugging Face
Code
github.com/deepseek-ai/deepseek-r1
Lab
DeepSeek

Changes

What changed
New paperAdded to the listSep 24, 2026

Sources: each lab's own papers and arXiv, with citation and upvote counts from Semantic Scholar and Hugging Face. One-line summaries are for orientation, not a substitute for the paper. Logos via logo.dev; trademarks belong to their owners.

New papers by email

Monday afternoons, only in weeks with new papers from the labs.

Double opt-in. Unsubscribe any time.