Foundational Automatic Evaluators: Scaling Multi-Task Generative Evaluator Training for Reasoning-Centric Domains research paper by Salesforce, 2025
Salesforce · Oct 20, 2025 · Reasoning · 4 upvotes 11 months ago
What it shows
FARE, a family of large-scale parameter evaluators, surpasses specialized RL-trained evaluators in both static benchmarks and real-world tasks through data-driven development and iterative rejection-sampling supervised finetuning.
Hugging Face's summary; not yet checked by hand.
More from Salesforce
All 33Other reasoning papers
TopicAbout this paper
- Authors
- Austin Xu, Xuan-Phi Nguyen, Yilun Zhou and 3 more
- arXiv
- 2510.17793 · PDF
- Citations
- Not counted yet · Semantic Scholar
- Upvotes
- 4 · Hugging Face
- Lab
- Salesforce · on Companies · on Acquisitions · on Quarterly · on Paydays · on TechConf
Changes
| What changed | |
|---|---|
| Sep 26, 2026today | New paperAdded to the listSep 26, 2026today |