3 papers
cs.CL2026
The Role of Prosodic and Lexical Cues in Turn-Taking with Self-Supervised Speech Representations
Sam OConnor Russell, Delphine Charuau, Naomi Harte
Fluid turn-taking remains a key challenge in human-robot interaction. Self-supervised speech representations (S3Rs) have driven many advances, but it remains unclear whether S3R-ba…
cs.SD2025
Visual Cues Support Robust Turn-taking Prediction in Noise
Sam O'Connor Russell, Naomi Harte
Accurate predictive turn-taking models (PTTMs) are essential for naturalistic human-robot interaction. However, little is known about their performance in noise. This study therefo…
cs.CL2025
Visual Cues Enhance Predictive Turn-Taking for Two-Party Human Interaction
Sam O'Connor Russell, Naomi Harte
Turn-taking is richly multimodal. Predictive turn-taking models (PTTMs) facilitate naturalistic human-robot interaction, yet most rely solely on speech. We introduce MM-VAP, a mult…