1 paper
A. Sayyad, J. Emmons, S. Jones +2
We report the empirical reliability of Gemini models as audio judges that score full-duplex agent conversations directly from the raw stereo waveform, tested across three models in…