collaborators
Showing eess.ASShow all

5 papers · 1 filter

eess.AS2025

Leveraging AM and FM Rhythm Spectrograms for Dementia Classification and Assessment

Parismita Gogoi, Vishwanath Pratap Singh, Seema Khadirnaikar +6

This study explores the potential of Rhythm Formant Analysis (RFA) to capture long-term temporal modulations in dementia speech. Specifically, we introduce RFA-derived rhythm spect…

eess.AS2025

Continuous Learning for Children's ASR: Overcoming Catastrophic Forgetting with Elastic Weight Consolidation and Synaptic Intelligence

Edem Ahadzi, Vishwanath Pratap Singh, Tomi Kinnunen +1

In this work, we present the first study addressing automatic speech recognition (ASR) for children in an online learning setting. This is particularly important for both child-cen…

eess.AS2025

ASVspoof 5: Design, Collection and Validation of Resources for Spoofing, Deepfake, and Adversarial Attack Detection Using Crowdsourced Speech

Xin Wang, Héctor Delgado, Hemlata Tak +26

ASVspoof 5 is the fifth edition in a series of challenges which promote the study of speech spoofing and deepfake attacks as well as the design of detection solutions. We introduce…

eess.AS2025

Causal Analysis of ASR Errors for Children: Quantifying the Impact of Physiological, Cognitive, and Extrinsic Factors

Vishwanath Pratap Singh, Md. Sahidullah, Tomi Kinnunen

The increasing use of children's automatic speech recognition (ASR) systems has spurred research efforts to improve the accuracy of models designed for children's speech in recent…

eess.AS2024

ROAR: Reinforcing Original to Augmented Data Ratio Dynamics for Wav2Vec2.0 Based ASR

Vishwanath Pratap Singh, Federico Malato, Ville Hautamaki +2

While automatic speech recognition (ASR) greatly benefits from data augmentation, the augmentation recipes themselves tend to be heuristic. In this paper, we address one of the heu…