12 citations · 30 across the 9 of their papers we have counts for
11 papers · 1 filter
Regularized Mask Tuning: Uncovering Hidden Knowledge in Pre-trained Vision-Language Models
Kecheng Zheng, Wei Wu, Ruili Feng +6
Prompt tuning and adapter tuning have shown great potential in transferring pre-trained vision-language models (VLMs) to various downstream tasks. In this work, we design a new typ…
Knowledge-Enhanced Hierarchical Information Correlation Learning for Multi-Modal Rumor Detection
Jiawei Liu, Jingyi Xie, Fanrui Zhang +2
The explosive growth of rumors with text and images on social media platforms has drawn great attention. Existing studies have made significant contributions to cross-modal informa…
Sounding Video Generator: A Unified Framework for Text-guided Sounding Video Generation
Jiawei Liu, Weining Wang, Sihan Chen +2
As a combination of visual and audio signals, video is inherently multi-modal. However, existing video generation methods are primarily intended for the synthesis of visual frames,…
Modality-Adaptive Mixup and Invariant Decomposition for RGB-Infrared Person Re-Identification
Zhipeng Huang, Jiawei Liu, Liang Li +2
RGB-infrared person re-identification is an emerging cross-modality re-identification task, which is very challenging due to significant modality discrepancy between RGB and infrar…
Debiased Batch Normalization via Gaussian Process for Generalizable Person Re-Identification
Jiawei Liu, Zhipeng Huang, Liang Li +2
Generalizable person re-identification aims to learn a model with only several labeled source domains that can perform well on unseen domains. Without access to the unseen domain,…
Pose-Guided Feature Learning with Knowledge Distillation for Occluded Person Re-Identification
Kecheng Zheng, Cuiling Lan, Wenjun Zeng +3
Occluded person re-identification (ReID) aims to match person images with occlusion. It is fundamentally challenging because of the serious occlusion which aggravates the misalignm…