14 citations · 35 across the 7 of their papers we have counts for
4 papers
Balanced Contrastive Learning for Long-Tailed Visual Recognition
Jianggang Zhu, Zheng Wang, Jingjing Chen +2
Real-world data typically follow a long-tailed distribution, where a few majority categories occupy most of the data while most minority categories contain a limited number of samp…
Unsupervised High-Resolution Portrait Gaze Correction and Animation
Jichao Zhang, Jingjing Chen, Hao Tang +5
This paper proposes a gaze correction and animation method for high-resolution, unconstrained portrait images, which can be trained without the gaze angle and the head pose annotat…
Unified Multimodal Pre-training and Prompt-based Tuning for Vision-Language Understanding and Generation
Tianyi Liu, Zuxuan Wu, Wenhan Xiong +2
Most existing vision-language pre-training methods focus on understanding tasks and use BERT-like objectives (masked language modeling and image-text matching) during pretraining.…
Cross-Modal Transferable Adversarial Attacks from Images to Videos
Zhipeng Wei, Jingjing Chen, Zuxuan Wu +1
Recent studies have shown that adversarial examples hand-crafted on one white-box model can be used to attack other black-box models. Such cross-model transferability makes it feas…