12 citations · 39 across the 11 of their papers we have counts for
Showing 2023 · cs.CVShow all
3 papers · 2 filters
cs.CV2023
Regularized Mask Tuning: Uncovering Hidden Knowledge in Pre-trained Vision-Language Models
Kecheng Zheng, Wei Wu, Ruili Feng +6
Prompt tuning and adapter tuning have shown great potential in transferring pre-trained vision-language models (VLMs) to various downstream tasks. In this work, we design a new typ…
cs.CV2023★ 2 cited
Knowledge-Enhanced Hierarchical Information Correlation Learning for Multi-Modal Rumor Detection
Jiawei Liu, Jingyi Xie, Fanrui Zhang +2
The explosive growth of rumors with text and images on social media platforms has drawn great attention. Existing studies have made significant contributions to cross-modal informa…
cs.CV2023★ 12 cited
Sounding Video Generator: A Unified Framework for Text-guided Sounding Video Generation
Jiawei Liu, Weining Wang, Sihan Chen +2
As a combination of visual and audio signals, video is inherently multi-modal. However, existing video generation methods are primarily intended for the synthesis of visual frames,…