No voice AI system dominates all capabilities; naturalness, expressiveness, identity stability, audio sensitivity, and transcription robustness vary independently, so voice AI should be evaluated as a multidimensional profile.
SD-Eval: A benchmark dataset for spoken dialogue understanding beyond words
1 Pith paper cite this work, alongside 7 external citations. Polarity classification is still indexing.
1
Pith paper citing it
7
external citations · OpenAlex
fields
cs.SD 1years
2026 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
RW-Voice-EQ Bench: A Real World Benchmark for Evaluating Voice AI Systems
No voice AI system dominates all capabilities; naturalness, expressiveness, identity stability, audio sensitivity, and transcription robustness vary independently, so voice AI should be evaluated as a multidimensional profile.