51 citations · 95 across the 4 of their papers we have counts for
9 papers
Cooperative Dual Attention for Audio-Visual Speech Enhancement with Facial Cues
Feixiang Wang, Shuang Yang, Shiguang Shan +1
In this work, we focus on leveraging facial cues beyond the lip region for robust Audio-Visual Speech Enhancement (AVSE). The facial region, encompassing the lip region, reflects a…
UniCon: Unified Context Network for Robust Active Speaker Detection
Yuanhang Zhang, Susan Liang, Shuang Yang +4
We introduce a new efficient framework, the Unified Context Network (UniCon), for robust active speaker detection (ASD). Traditional methods for ASD usually operate on each candida…
Learn an Effective Lip Reading Model without Pains
Dalu Feng, Shuang Yang, Shiguang Shan +1
Lip reading, also known as visual speech recognition, aims to recognize the speech content from videos by analyzing the lip dynamics. There have been several appealing progress in…
Synchronous Bidirectional Learning for Multilingual Lip Reading
Mingshuang Luo, Shuang Yang, Xilin Chen +2
Lip reading has received increasing attention in recent years. This paper focuses on the synergy of multilingual lip reading. There are about as many as 7000 languages in the world…
Mutual Information Maximization for Effective Lip Reading
Xing Zhao, Shuang Yang, Shiguang Shan +1
Lip reading has received an increasing research interest in recent years due to the rapid development of deep learning and its widespread potential applications. One key point to o…
Deformation Flow Based Two-Stream Network for Lip Reading
Jingyun Xiao, Shuang Yang, Yuanhang Zhang +2
Lip reading is the task of recognizing the speech content by analyzing movements in the lip region when people are speaking. Observing on the continuity in adjacent frames in the s…