Training language models to follow instructions with human feedback research paper by OpenAI, 2022
OpenAI · Mar 4, 2022 · Alignment and safety · 24,241 citations · 26 upvotes
What it shows
InstructGPT: fine-tuning on human feedback made a 1.3B model preferred over the 175B GPT-3.
Summarised by hand from the abstract.
More from OpenAI
All 5Other alignment and safety papers
TopicAbout this paper
- Authors
- Long Ouyang, Jeff Wu, Xu Jiang and 17 more
- arXiv
- 2203.02155 · PDF
- Venue
- Neural Information Processing Systems
- Citations
- 24,241, 2,459 influential · Semantic Scholar
- Upvotes
- 26 · Hugging Face
- Code
- github.com/openai/following-instructions-human-feedback
- Lab
- OpenAI · on Companies · on Acquisitions · on Paydays · on Releases · on TechConf
Changes
| What changed | |
|---|---|
| Sep 24, 2026 | New paperAdded to the listSep 24, 2026 |
Sources: each lab's own papers and arXiv, with citation and upvote counts from Semantic Scholar and Hugging Face. One-line summaries are for orientation, not a substitute for the paper. Logos via logo.dev; trademarks belong to their owners.