Why Language Models Hallucinate research paper by OpenAI, 2025
OpenAI · Sep 4, 2025 · Alignment and safety · 329 citations · 200 upvotes
What it shows
Models hallucinate because training and benchmarks reward confident guessing over saying they do not know.
Summarised by hand from the abstract.
More from OpenAI
All 5Other alignment and safety papers
TopicAbout this paper
- Authors
- Adam Tauman Kalai, Ofir Nachum, Santosh S. Vempala and 1 more
- arXiv
- 2509.04664 · PDF
- Venue
- arXiv.org
- Citations
- 329, 19 influential · Semantic Scholar
- Upvotes
- 200 · Hugging Face
- Lab
- OpenAI · on Companies · on Acquisitions · on Paydays · on Releases · on TechConf
Changes
| What changed | |
|---|---|
| Sep 24, 2026 | New paperAdded to the listSep 24, 2026 |
Sources: each lab's own papers and arXiv, with citation and upvote counts from Semantic Scholar and Hugging Face. One-line summaries are for orientation, not a substitute for the paper. Logos via logo.dev; trademarks belong to their owners.