25 citations · 73 across the 18 of their papers we have counts for
18 papers
WRF4CIR: Weight-Regularized Fine-Tuning Network for Composed Image Retrieval
Yizhuo Xu, Chaojian Yu, Yuanjie Shao +3
Composed Image Retrieval (CIR) task aims to retrieve target images based on reference images and modification texts. Current CIR methods primarily rely on fine-tuning vision-langua…
Adjustable Text-Guided Backdoor Attacks with Natural-Word Triggers on Multimodal Pretrained Models
Yiyang Zhang, Chaojian Yu, Ziming Hong +4
This paper presents Text-Guided Backdoor (TGB), an adjustable backdoor attack against multimodal pretrained models that uses natural-word triggers, namely words that can naturally…
Learning Unpaired Image Dehazing with Physics-based Rehazy Generation
Haoyou Deng, Zhiqiang Li, Feng Zhang +6
Overfitting to synthetic training pairs remains a critical challenge in image dehazing, leading to poor generalization capability to real-world scenarios. To address this issue, ex…
NTIRE 2025 Challenge on Cross-Domain Few-Shot Object Detection: Methods and Results
Yuqian Fu, Xingyu Qiu, Bin Ren +59
Cross-Domain Few-Shot Object Detection (CD-FSOD) poses significant challenges to existing object detection and few-shot detection models when applied across domains. In conjunction…
Object-Aware Video Matting with Cross-Frame Guidance
Huayu Zhang, Dongyue Wu, Yuanjie Shao +2
Recently, trimap-free methods have drawn increasing attention in human video matting due to their promising performance. Nevertheless, these methods still suffer from the lack of d…
CTR-Driven Advertising Image Generation with Multimodal Large Language Models
Xingye Chen, Wei Feng, Zhenbang Du +16
In web data, advertising images are crucial for capturing user attention and improving advertising effectiveness. Most existing methods generate background for products primarily f…