ResearchRubrics: A Benchmark of Prompts and Rubrics For Evaluating Deep Research Agents research paper by Scale AI, 2025
Scale AI · Nov 10, 2025 · Agents and evaluation · 10 upvotes 10 months ago
What it shows
ResearchRubrics is a benchmark for evaluating deep research agents, using expert rubrics to assess their factual grounding, reasoning, and clarity across diverse, complex tasks.
Hugging Face's summary; not yet checked by hand.
More from Scale AI
All 25Other agents and evaluation papers
TopicAbout this paper
- Authors
- Manasi Sharma, Chen Bo Calvin Zhang, Chaithanya Bandi and 13 more
- arXiv
- 2511.07685 · PDF
- Citations
- Not counted yet · Semantic Scholar
- Upvotes
- 10 · Hugging Face
- Lab
- Scale AI · on Companies · on Acquisitions
Changes
| What changed | |
|---|---|
| Sep 26, 2026today | New paperAdded to the listSep 26, 2026today |