activity
20162022
most citedAdvances in Online Audio-Visual Meeting Transcription

17 citations · 105 across the 35 of their papers we have counts for

collaborators

49 papers

eess.AS2022

Speech separation with large-scale self-supervised learning

Zhuo Chen, Naoyuki Kanda, Jian Wu +6

Self-supervised learning (SSL) methods such as WavLM have shown promising speech separation (SS) results in small-scale simulation-based experiments. In this work, we extend the ex…

eess.AS2022

Exploring WavLM on Speech Enhancement

Hyungchan Song, Sanyuan Chen, Zhuo Chen +5

There is a surge in interest in self-supervised learning approaches for end-to-end speech encoding in recent years as they have achieved great success. Especially, WavLM showed sta…

eess.AS20221 cited

Simulating realistic speech overlaps improves multi-talker ASR

Muqiao Yang, Naoyuki Kanda, Xiaofei Wang +5

Multi-talker automatic speech recognition (ASR) has been studied to generate transcriptions of natural conversation including overlapping speech of multiple speakers. Due to the di…

eess.AS2022

An Adapter based Multi-label Pre-training for Speech Separation and Enhancement

Tianrui Wang, Xie Chen, Zhuo Chen +2

In recent years, self-supervised learning (SSL) has achieved tremendous success in various speech tasks due to its power to extract representations from massive unlabeled data. How…

eess.AS2022

Self-supervised learning with bi-label masked speech prediction for streaming multi-talker speech recognition

Zili Huang, Zhuo Chen, Naoyuki Kanda +6

Self-supervised learning (SSL), which utilizes the input data itself for representation learning, has achieved state-of-the-art results for various downstream speech tasks. However…

cs.CL2022

Real-time Speech Interruption Analysis: From Cloud to Client Deployment

Quchen Fu, Szu-Wei Fu, Yaran Fan +4

Meetings are an essential form of communication for all types of organizations, and remote collaboration systems have been much more widely used since the COVID-19 pandemic. One ma…