1 citations · 1 across the 2 of their papers we have counts for
3 papers
cs.CV2025
LipGen: Viseme-Guided Lip Video Generation for Enhancing Visual Speech Recognition
Bowen Hao, Dongliang Zhou, Xiaojie Li +4
Visual speech recognition (VSR), commonly known as lip reading, has garnered significant attention due to its wide-ranging practical applications. The advent of deep learning techn…
cs.SD2025
AVE Speech: A Comprehensive Multi-Modal Dataset for Speech Recognition Integrating Audio, Visual, and Electromyographic Signals
Dongliang Zhou, Yakun Zhang, Jinghan Wu +3
The global aging population faces considerable challenges, particularly in communication, due to the prevalence of hearing and speech impairments. To address these, we introduce th…
cs.AI2024★ 1 cited
Landmark-Guided Cross-Speaker Lip Reading with Mutual Information Regularization
Linzhi Wu, Xingyu Zhang, Yakun Zhang +5
Lip reading, the process of interpreting silent speech from visual lip movements, has gained rising attention for its wide range of realistic applications. Deep learning approaches…