activity
20192023
most citedDeep Audio-Visual Learning: A Survey

16 citations · 17 across the 4 of their papers we have counts for

collaborators
Showing cs.CVShow all

5 papers · 1 filter

cs.CV2023

Learning Cross-modality Information Bottleneck Representation for Heterogeneous Person Re-Identification

Haichao Shi, Mandi Luo, Xiao-Yu Zhang +1

Visible-Infrared person re-identification (VI-ReID) is an important and challenging task in intelligent video surveillance. Existing methods mainly focus on learning a shared featu…

cs.CV20221 cited

ScoreMix: A Scalable Augmentation Strategy for Training GANs with Limited Data

Jie Cao, Mandi Luo, Junchi Yu +2

Generative Adversarial Networks (GANs) typically suffer from overfitting when limited training data is available. To facilitate GAN training, current methods propose to use data-sp…

cs.CV2020

Unsupervised Contrastive Photo-to-Caricature Translation based on Auto-distortion

Yuhe Ding, Xin Ma, Mandi Luo +2

Photo-to-caricature translation aims to synthesize the caricature as a rendered image exaggerating the features through sketching, pencil strokes, or other artistic drawings. Style…

cs.CV202016 cited

Deep Audio-Visual Learning: A Survey

Hao Zhu, Mandi Luo, Rui Wang +2

Audio-visual learning, aimed at exploiting the relationship between audio and visual modalities, has drawn considerable attention since deep learning started to be used successfull…

cs.CV2019

Exploiting Style and Attention in Real-World Super-Resolution

Xin Ma, Yi Li, Huaibo Huang +2

Real-world image super-resolution (SR) is a challenging image translation problem. Low-resolution (LR) images are often generated by various unknown transformations rather than by…