SciPredict: Can LLMs Predict the Outcomes of Scientific Experiments in Natural Sciences? research paper by Scale AI, 2026
Scale AI · Apr 12, 2026 · Applied AI · 4 upvotes 5 months ago
What it shows
SciPredict benchmark reveals that large language models struggle to accurately predict scientific experiment outcomes and cannot reliably assess prediction confidence, unlike human experts who show better calibration and performance when experiments are deemed predictable.
Hugging Face's summary; not yet checked by hand.
More from Scale AI
All 25Other applied ai papers
TopicAbout this paper
- Authors
- Udari Madhushani Sehwag, Elaine Lau, Haniyeh Ehsani Oskouie and 14 more
- arXiv
- 2604.10718 · PDF
- Citations
- Not counted yet · Semantic Scholar
- Upvotes
- 4 · Hugging Face
- Lab
- Scale AI · on Companies · on Acquisitions
Changes
| What changed | |
|---|---|
| Sep 26, 2026today | New paperAdded to the listSep 26, 2026today |