OpenAI o1 System Card research paper by OpenAI, 2024
OpenAI · Dec 21, 2024 · Alignment and safety · 0 citations · 38 upvotes · unverified
What it shows
The o1 model series, trained with reinforcement learning and chain of thought, enhances safety and robustness by reasoning about policies, leading to superior performance on risk benchmarks while highlighting the need for robust alignment and risk management.
UnverifiedHugging Face's summary; not yet checked by hand.
More from OpenAI
All 8Other alignment and safety papers
TopicAbout this paper
- Authors
- OpenAI, Aaron Jaech, Adam Kalai and 262 more
- arXiv
- 2412.16720 · PDF
- Citations
- 0, 0 influential · Semantic Scholar
- Upvotes
- 38 · Hugging Face
- Lab
- OpenAI · on Companies · on Acquisitions · on Paydays · on Releases · on TechConf
Changes
| What changed | |
|---|---|
| Sep 25, 2026 | Influential citationsfirst count: 0Sep 25, 2026 |
| Sep 25, 2026 | Citationsfirst count: 0Sep 25, 2026 |
| Sep 25, 2026 | New paperFound by the weekly scan, unverifiedSep 25, 2026 |
Sources: each lab's own papers and arXiv, with citation and upvote counts from Semantic Scholar and Hugging Face. One-line summaries are for orientation, not a substitute for the paper. Logos via logo.dev; trademarks belong to their owners.