Showing cs.SDShow all
2 papers · 1 filter
cs.SD2025
On Barriers to Archival Audio Processing
Peter Sullivan, Muhammad Abdul-Mageed
In this study, we leverage a unique UNESCO collection of mid-20th century radio recordings to probe the robustness of modern off-the-shelf language identification (LID) and speaker…
cs.SD2024
What Does it Take to Generalize SER Model Across Datasets? A Comprehensive Benchmark
Adham Ibrahim, Shady Shehata, Ajinkya Kulkarni +2
Speech emotion recognition (SER) is essential for enhancing human-computer interaction in speech-based applications. Despite improvements in specific emotional datasets, there is s…