Showing cs.SDShow all
2 papers · 1 filter
cs.SD2026
voice2mode: Phonation Mode Classification in Singing using Self-Supervised Speech Models
Aju Ani Justus, Ruchit Agrawal, Sudarsana Reddy Kadiri +1
We present voice2mode, a method for classification of four singing phonation modes (breathy, neutral (modal), flow, and pressed) using embeddings extracted from large self-supervis…
cs.SD2024
MMSD-Net: Towards Multi-modal Stuttering Detection
Liangyu Nie, Sudarsana Reddy Kadiri, Ruchit Agrawal
Stuttering is a common speech impediment that is caused by irregular disruptions in speech production, affecting over 70 million people across the world. Standard automatic speech…