TutorBench: A Benchmark To Assess Tutoring Capabilities Of Large Language Models research paper by Scale AI, 2025
Scale AI · Oct 3, 2025 · Agents and evaluation · 2 upvotes 11 months ago
What it shows
TutorBench is a dataset and benchmark for evaluating the tutoring skills of large language models, showing significant room for improvement in adaptive explanations, feedback, and active learning.
Hugging Face's summary; not yet checked by hand.
More from Scale AI
All 25Other agents and evaluation papers
TopicAbout this paper
- Authors
- Rakshith S Srinivasa, Zora Che, Chen Bo Calvin Zhang and 11 more
- arXiv
- 2510.02663 · PDF
- Citations
- Not counted yet · Semantic Scholar
- Upvotes
- 2 · Hugging Face
- Lab
- Scale AI · on Companies · on Acquisitions
Changes
| What changed | |
|---|---|
| Sep 26, 2026today | New paperAdded to the listSep 26, 2026today |