3 papers
cs.CL2025
Pitch Accent Detection improves Pretrained Automatic Speech Recognition
David Sasu, Natalie Schluter
We show the performance of Automatic Speech Recognition (ASR) systems that use semi-supervised speech representations can be boosted by a complimentary pitch accent detection modul…
cs.CL2025
Akan Cinematic Emotions (ACE): A Multimodal Multi-party Dataset for Emotion Recognition in Movie Dialogues
David Sasu, Zehui Wu, Ziwei Gong +5
In this paper, we introduce the Akan Conversation Emotion (ACE) dataset, the first multimodal emotion dialogue dataset for an African language, addressing the significant lack of r…
cs.RO2025
Enhancing Speech Instruction Understanding and Disambiguation in Robotics via Speech Prosody
David Sasu, Kweku Andoh Yamoah, Benedict Quartey +1
Enabling robots to accurately interpret and execute spoken language instructions is essential for effective human-robot collaboration. Traditional methods rely on speech recognitio…