activity
20182023
most citedLearn an Effective Lip Reading Model without Pains

51 citations · 95 across the 4 of their papers we have counts for

collaborators

9 papers

cs.CV2023

Cooperative Dual Attention for Audio-Visual Speech Enhancement with Facial Cues

Feixiang Wang, Shuang Yang, Shiguang Shan +1

In this work, we focus on leveraging facial cues beyond the lip region for robust Audio-Visual Speech Enhancement (AVSE). The facial region, encompassing the lip region, reflects a…

cs.CV202139 cited

UniCon: Unified Context Network for Robust Active Speaker Detection

Yuanhang Zhang, Susan Liang, Shuang Yang +4

We introduce a new efficient framework, the Unified Context Network (UniCon), for robust active speaker detection (ASD). Traditional methods for ASD usually operate on each candida…

cs.CV202051 cited

Learn an Effective Lip Reading Model without Pains

Dalu Feng, Shuang Yang, Shiguang Shan +1

Lip reading, also known as visual speech recognition, aims to recognize the speech content from videos by analyzing the lip dynamics. There have been several appealing progress in…

cs.CV20205 cited

Synchronous Bidirectional Learning for Multilingual Lip Reading

Mingshuang Luo, Shuang Yang, Xilin Chen +2

Lip reading has received increasing attention in recent years. This paper focuses on the synergy of multilingual lip reading. There are about as many as 7000 languages in the world…

cs.CV2020

Mutual Information Maximization for Effective Lip Reading

Xing Zhao, Shuang Yang, Shiguang Shan +1

Lip reading has received an increasing research interest in recent years due to the rapid development of deep learning and its widespread potential applications. One key point to o…

cs.CV2020

Deformation Flow Based Two-Stream Network for Lip Reading

Jingyun Xiao, Shuang Yang, Yuanhang Zhang +2

Lip reading is the task of recognizing the speech content by analyzing movements in the lip region when people are speaking. Observing on the continuity in adjacent frames in the s…