Skip to content
AI Atlas
PaperActive

EVA-Bench: A New End-to-end Framework for Evaluating Voice Agents

arxiv.org/abs/2605.13841

quality89

Updated 2 h ago · first seen 11 Sept 2026

paper_01M294GQ1JB1N52VCRFXHN0Y3Q

Published
11 Sept 2026
T1 · 2 h ago
arXiv
2605.13841
T1 · 2 h ago
Category
cs.SD
T1 · 2 h ago

As of

Rewind the record: see this entity's attributes exactly as AI Atlas knew them on a given day.

Claim history · Authors

1 claims · 1 propertiesShow all properties

Authorsauthors1

Claim history for Authors
ValueValid from → toStatusSourceConfidenceExtractor
Tara Bogavelli, Gabrielle Gauthier Melan\c{c}on, Katrina Stankiewicz, Oluwanifemi Bamgbose, Fanny Riols, Hoang H. Nguyen, Raghav Mehndiratta, Lindsay Devon Brin, Joseph Marinier, Hari Subramani, Anil Madamala, Sridhar Krishna Nemala, Srinivas SunkaracurrentcurrentarXiv (Atom API + RSS)T1highdeterministic

Claims are temporal and append-only: a new observation closes the previous claim (valid_to) instead of overwriting it. Conflicting claims from different sources are kept side by side and flagged — never averaged. Methodology →