DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning research paper by DeepSeek, 2025
DeepSeek · Jan 22, 2025 · Reasoning · 5,691 citations · 462 upvotes
What it shows
Reinforcement learning taught a model long step-by-step reasoning on par with OpenAI o1.
Summarised by hand from the abstract.
More from DeepSeek
All 2Other reasoning papers
TopicAbout this paper
- Authors
- DeepSeek-AI, Daya Guo, Dejian Yang and 197 more
- arXiv
- 2501.12948 · PDF
- Venue
- Nature
- Citations
- 5,691, 921 influential · Semantic Scholar
- Upvotes
- 462 · Hugging Face
- Code
- github.com/deepseek-ai/deepseek-r1
- Lab
- DeepSeek
Changes
| What changed | |
|---|---|
| Sep 24, 2026 | New paperAdded to the listSep 24, 2026 |
Sources: each lab's own papers and arXiv, with citation and upvote counts from Semantic Scholar and Hugging Face. One-line summaries are for orientation, not a substitute for the paper. Logos via logo.dev; trademarks belong to their owners.