Skip to content
Papers.

Training language models to follow instructions with human feedback research paper by OpenAI, 2022

OpenAI · Mar 4, 2022 · Alignment and safety · 24,241 citations · 26 upvotes

Read on arXiv

What it shows

InstructGPT: fine-tuning on human feedback made a 1.3B model preferred over the 175B GPT-3.

Summarised by hand from the abstract.

More from OpenAI

All 5
Topic
About this paper
Authors
Long Ouyang, Jeff Wu, Xu Jiang and 17 more
arXiv
2203.02155 · PDF
Venue
Neural Information Processing Systems
Citations
24,241, 2,459 influential · Semantic Scholar
Upvotes
26 · Hugging Face
Code
github.com/openai/following-instructions-human-feedback
Lab
OpenAI · on Companies · on Acquisitions · on Paydays · on Releases · on TechConf

Changes

What changed
New paperAdded to the listSep 24, 2026

Sources: each lab's own papers and arXiv, with citation and upvote counts from Semantic Scholar and Hugging Face. One-line summaries are for orientation, not a substitute for the paper. Logos via logo.dev; trademarks belong to their owners.

New papers by email

Monday afternoons, only in weeks with new papers from the labs.

Double opt-in. Unsubscribe any time.