EVA-Bench: A New End-to-end Framework for Evaluating Voice Agents research paper by ServiceNow, 2026
ServiceNow · May 13, 2026 · Agents and evaluation · 77 upvotes · unverified 4 months ago
What it shows
EVA-Bench presents a comprehensive evaluation framework for voice agents that simulates realistic conversations and measures performance across multiple voice-specific failure modes using novel accuracy and experience metrics.
UnverifiedHugging Face's summary; not yet checked by hand.
More from ServiceNow
All 20Other agents and evaluation papers
TopicAbout this paper
- Authors
- Tara Bogavelli, Gabrielle Gauthier Melançon, Katrina Stankiewicz and 10 more
- arXiv
- 2605.13841 · PDF
- Citations
- Not counted yet · Semantic Scholar
- Upvotes
- 77 · Hugging Face
- Code
- github.com/ServiceNow/eva
- Lab
- ServiceNow · on Companies · on Acquisitions · on Quarterly · on Paydays · on TechConf
Changes
| What changed | |
|---|---|
| Sep 26, 2026today | New paperFound by the weekly scan, unverifiedSep 26, 2026today |