1 citations · 1 across the 2 of their papers we have counts for
2 papers
eess.AS2024
The USTC-NERCSLIP Systems for the CHiME-8 MMCSG Challenge
Ya Jiang, Hongbo Lan, Jun Du +2
In the two-person conversation scenario with one wearing smart glasses, transcribing and displaying the speaker's content in real-time is an intriguing application, providing a pri…
eess.AS2023★ 1 cited
The Multimodal Information Based Speech Processing (MISP) 2023 Challenge: Audio-Visual Target Speaker Extraction
Shilong Wu, Chenxi Wang, Hang Chen +13
Previous Multimodal Information based Speech Processing (MISP) challenges mainly focused on audio-visual speech recognition (AVSR) with commendable success. However, the most advan…