activity
20192022
most citedUnsupervised Cross-Modal Distillation for Thermal Infrared Tracking

31 citations · 46 across the 4 of their papers we have counts for

collaborators

5 papers

cs.CV20224 cited

An Audio-Visual Attention Based Multimodal Network for Fake Talking Face Videos Detection

Ganglai Wang, Peng Zhang, Lei Xie +3

DeepFake based digital facial forgery is threatening the public media security, especially when lip manipulation has been used in talking face generation, the difficulty of fake vi…

cs.CV20229 cited

Attention-Based Lip Audio-Visual Synthesis for Talking Face Generation in the Wild

Ganglai Wang, Peng Zhang, Lei Xie +2

Talking face generation with great practical significance has attracted more attention in recent audio-visual studies. How to achieve accurate lip synchronization is a long-standin…

cs.SD20222 cited

Audio-visual speech separation based on joint feature representation with cross-modal attention

Junwen Xiong, Peng Zhang, Lei Xie +3

Multi-modal based speech separation has exhibited a specific advantage on isolating the target character in multi-talker noisy environments. Unfortunately, most of current separati…

cs.CV202131 cited

Unsupervised Cross-Modal Distillation for Thermal Infrared Tracking

Jingxian Sun, Lichao Zhang, Yufei Zha +4

The target representation learned by convolutional neural networks plays an important role in Thermal Infrared (TIR) tracking. Currently, most of the top-performing TIR trackers ar…

cs.CV2019

Push for Quantization: Deep Fisher Hashing

Yunqiang Li, Wenjie Pei, Yufei zha +1

Current massive datasets demand light-weight access for analysis. Discrete hashing methods are thus beneficial because they map high-dimensional data to compact binary codes that a…