97 citations · 97 across the 1 of their papers we have counts for
4 papers · 1 filter
Listening to Multi-talker Conversations: Modular and End-to-end Perspectives
Desh Raj
Since the first speech recognition systems were built more than 30 years ago, improvement in voice technology has enabled applications such as smart assistants and automated custom…
On Speaker Attribution with SURT
Desh Raj, Matthew Wiesner, Matthew Maciejewski +3
The Streaming Unmixing and Recognition Transducer (SURT) has recently become a popular framework for continuous, streaming, multi-talker speech recognition (ASR). With advances in…
Learning from Flawed Data: Weakly Supervised Automatic Speech Recognition
Dongji Gao, Hainan Xu, Desh Raj +3
Training automatic speech recognition (ASR) systems requires large amounts of well-curated paired data. However, human annotators usually perform "non-verbatim" transcription, whic…
The CHiME-7 DASR Challenge: Distant Meeting Transcription with Multiple Devices in Diverse Scenarios
Samuele Cornell, Matthew Wiesner, Shinji Watanabe +8
The CHiME challenges have played a significant role in the development and evaluation of robust automatic speech recognition (ASR) systems. We introduce the CHiME-7 distant ASR (DA…