4 citations · 10 across the 3 of their papers we have counts for
2 papers
cs.CV2022★ 4 cited
An Audio-Visual Attention Based Multimodal Network for Fake Talking Face Videos Detection
Ganglai Wang, Peng Zhang, Lei Xie +3
DeepFake based digital facial forgery is threatening the public media security, especially when lip manipulation has been used in talking face generation, the difficulty of fake vi…
cs.SD2022★ 2 cited
Audio-visual speech separation based on joint feature representation with cross-modal attention
Junwen Xiong, Peng Zhang, Lei Xie +3
Multi-modal based speech separation has exhibited a specific advantage on isolating the target character in multi-talker noisy environments. Unfortunately, most of current separati…