MoReBench: Evaluating Procedural and Pluralistic Moral Reasoning in Language Models, More than Outcomes research paper by Scale AI, 2025
Scale AI · Oct 18, 2025 · Reasoning · 2 upvotes 11 months ago
What it shows
MoReBench and MoReBench-Theory provide benchmarks for evaluating AI's moral reasoning and decision-making processes, highlighting the need for process-focused evaluation and transparency in AI systems.
Hugging Face's summary; not yet checked by hand.
More from Scale AI
All 25Other reasoning papers
TopicAbout this paper
- Authors
- Yu Ying Chiu, Michael S. Lee, Rachel Calcott and 15 more
- arXiv
- 2510.16380 · PDF
- Citations
- Not counted yet · Semantic Scholar
- Upvotes
- 2 · Hugging Face
- Code
- github.com/morebench/morebench
- Lab
- Scale AI · on Companies · on Acquisitions
Changes
| What changed | |
|---|---|
| Sep 26, 2026today | New paperAdded to the listSep 26, 2026today |