3 papers
cs.SD2025
Data-independent Beamforming for End-to-end Multichannel Multi-speaker ASR
Can Cui, Paul Magron, Mostafa Sadeghi +1
Automatic speech recognition (ASR) in multichannel, multi-speaker scenarios remains challenging due to ambient noise, reverberation and overlapping speakers. In this paper, we prop…
cs.CL2025
End-to-end Joint Punctuated and Normalized ASR with a Limited Amount of Punctuated Training Data
Can Cui, Imran Ahamad Sheikh, Mostafa Sadeghi +1
Joint punctuated and normalized automatic speech recognition (ASR) aims at outputing transcripts with and without punctuation and casing. This task remains challenging due to the l…
cs.CL2025
Joint Beamforming and Speaker-Attributed ASR for Real Distant-Microphone Meeting Transcription
Can Cui, Imran Ahamad Sheikh, Mostafa Sadeghi +1
Distant-microphone meeting transcription is a challenging task. State-of-the-art end-to-end speaker-attributed automatic speech recognition (SA-ASR) architectures lack a multichann…