activity
20192022
most citedDeep Audio-Visual Learning: A Survey

16 citations · 27 across the 6 of their papers we have counts for

collaborators

6 papers

cs.CV20226 cited

ProxyMix: Proxy-based Mixup Training with Label Refinery for Source-Free Domain Adaptation

Yuhe Ding, Lijun Sheng, Jian Liang +2

Unsupervised domain adaptation (UDA) aims to transfer knowledge from a labeled source domain to an unlabeled target domain. Owing to privacy concerns and heavy data transmission, s…

cs.CV20204 cited

Viewpoint-aware Progressive Clustering for Unsupervised Vehicle Re-identification

Aihua Zheng, Xia Sun, Chenglong Li +1

Vehicle re-identification (Re-ID) is an active task due to its importance in large-scale intelligent monitoring in smart cities. Despite the rapid progress in recent years, most ex…

cs.CV2020

Unsupervised Contrastive Photo-to-Caricature Translation based on Auto-distortion

Yuhe Ding, Xin Ma, Mandi Luo +2

Photo-to-caricature translation aims to synthesize the caricature as a rendered image exaggerating the features through sketching, pencil strokes, or other artistic drawings. Style…

cs.CV20201 cited

Lets Play Music: Audio-driven Performance Video Generation

Hao Zhu, Yi Li, Feixia Zhu +2

We propose a new task named Audio-driven Per-formance Video Generation (APVG), which aims to synthesizethe video of a person playing a certain instrument guided bya given music aud…

cs.CV202016 cited

Deep Audio-Visual Learning: A Survey

Hao Zhu, Mandi Luo, Rui Wang +2

Audio-visual learning, aimed at exploiting the relationship between audio and visual modalities, has drawn considerable attention since deep learning started to be used successfull…

cs.CV2019

Multi-Adapter RGBT Tracking

Chenglong Li, Andong Lu, Aihua Zheng +2

The task of RGBT tracking aims to take the complementary advantages from visible spectrum and thermal infrared data to achieve robust visual tracking, and receives more and more at…