collaborators

17 papers

eess.AS2026

Navigating Speech Enhancement for Real-Time MRI: A Systematic Assessment of Signal Quality, Source Preservation, and Downstream Tasks

Huang-Cheng Chou, Sean Foley, Haley Hsu +10

Audio recorded during real-time magnetic resonance imaging (rtMRI) is heavily contaminated by scanner noise, but it remains unclear whether general-purpose speech enhancement impro…

eess.AS2026

Layer-wise Cross-Lingual Depression Detection from Speech: Analysis with Contrastive Alignment

Anisha Pattanayak, Hanie Kang, Huang-Cheng Chou +2

Significant disparities exist in the diagnosis and clinical presentation of depression across different linguistic populations. Speech-based depression detection performs well mono…

eess.AS2026

Autoencoder based optimized SSL representations: Complexity Minimization and improved Dysarthric ASR

Paban Sapkota, Hemant Kumar Kathania, Mikko Kurimo +2

Self-supervised learning (SSL) models extract rich speech representations but often come with high-dimensional features, increasing computational complexity. This work explores an…

eess.AS2026

An Approach to Simultaneous Acquisition of Real-Time MRI Video, EEG, and Surface EMG for Articulatory, Brain, and Muscle Activity During Speech Production

Jihwan Lee, Parsa Razmara, Kevin Huang +16

Speech production is a complex process spanning neural planning, motor control, muscle activation, and articulatory kinematics. While the acoustic speech signal is the most accessi…

eess.AS2026

Improving End-to-End Speech Recognition for Dysarthric Speech through In-Domain Data Augmentation

Paban Sapkota, Hemant Kumar Kathania, Sudarsana Reddy Kadiri +1

Dysarthric speech recognition is crucial for facilitating effective communication among individuals with dysarthria. However, accurately recognizing dysarthric speech poses signifi…

eess.AS2026

Systematic Study of Dysarthric Speech Recognition: Spectral Features and Acoustic Models

Paban Sapkota, Hemant Kumar Kathania, Mikko Kurimo +2

The challenge associated with recognizing dysarthric speech primarily arises from pronounced acoustic variability attributed to impaired articulatory precision. Past research has d…