5 citations · 10 across the 7 of their papers we have counts for
5 papers · 1 filter
SafeNexus: Discovering and Steering Modality-Universal Safety Neurons in MLLMs
Jian Yu, Fei Shen, Cong Wang +5
Although Large Language Models (LLMs) have demonstrated promising safety performance, extending them to Multimodal Large Language Models (MLLMs) exposes a significant gap between e…
LeCoT: revisiting network architecture for two-view correspondence pruning
Luanyuan Dai, Xiaoyu Du, Jinhui Tang
Two-view correspondence pruning aims to accurately remove incorrect correspondences (outliers) from initial ones and is widely applied to various computer vision tasks. Current pop…
TEST-V: TEst-time Support-set Tuning for Zero-shot Video Classification
Rui Yan, Jin Wang, Hongyu Qu +4
Recently, adapting Vision Language Models (VLMs) to zero-shot visual classification by tuning class embedding with a few prompts (Test-time Prompt Tuning, TPT) or replacing class n…
MGNet: Learning Correspondences via Multiple Graphs
Luanyuan Dai, Xiaoyu Du, Hanwang Zhang +1
Learning correspondences aims to find correct correspondences (inliers) from the initial correspondence set with an uneven correspondence distribution and a low inlier rate, which…
BiSTNet: Semantic Image Prior Guided Bidirectional Temporal Feature Fusion for Deep Exemplar-based Video Colorization
Yixin Yang, Zhongzheng Peng, Xiaoyu Du +3
How to effectively explore the colors of reference exemplars and propagate them to colorize each frame is vital for exemplar-based video colorization. In this paper, we present an…