evaluation metrics 1online streaming 1real-world recordings 1speech separation 1target speaker extraction 1
From the 1 of 6 linked papers with an AI index.
Showing cs.SDShow all
2 papers · 1 filter
cs.SD2025
Audio-Visual Speech Enhancement In Complex Scenarios With Separation And Dereverberation Joint Modeling
Jiarong Du, Zhan Jin, Peijun Yang +4
Audio-visual speech enhancement (AVSE) is a task that uses visual auxiliary information to extract a target speaker's speech from mixed audio. In real-world scenarios, there often…
cs.SD2024
AS-70: A Mandarin stuttered speech dataset for automatic speech recognition and stuttering event detection
Rong Gong, Hongfei Xue, Lezhi Wang +11
The rapid advancements in speech technologies over the past two decades have led to human-level performance in tasks like automatic speech recognition (ASR) for fluent speech. Howe…