1 paper
David Ayllon, Alice Baird, Jeffrey Brooks +11
Current voice AI benchmarks typically evaluate isolated capabilities such as speech intelligibility, word error rate, or text-based dialogue quality, but they rarely test whether s…