M-RewardBench: Evaluating Reward Models in Multilingual Settings research paper by Cohere, 2024
Cohere · Oct 20, 2024 · Agents and evaluation · 10 upvotes 1 year ago
What it shows
A systematic evaluation of reward models in multilingual settings reveals significant performance gaps and dependencies on translation quality and resource availability.
Hugging Face's summary; not yet checked by hand.
More from Cohere
All 30Other agents and evaluation papers
TopicAbout this paper
- Authors
- Srishti Gureja, Lester James V. Miranda, Shayekh Bin Islam and 7 more
- arXiv
- 2410.15522 · PDF
- Citations
- Not counted yet · Semantic Scholar
- Upvotes
- 10 · Hugging Face
- Code
- github.com/Cohere-Labs-Community/m-rewardbench
- Lab
- Cohere · on Companies · on Acquisitions · on Releases
Changes
| What changed | |
|---|---|
| Sep 26, 2026today | New paperAdded to the listSep 26, 2026today |