3 citations · 3 across the 1 of their papers we have counts for
3 papers
cs.CV2026★ 3 cited
CODER: Coupled Diversity-Sensitive Momentum Contrastive Learning for Image-Text Retrieval
Haoran Wang, Dongliang He, Wenhao Wu +7
Image-Text Retrieval (ITR) is challenging in bridging visual and lingual modalities. Contrastive learning has been adopted by most prior arts. Except for limited amount of negative…
cs.CV2025
DeltaEdit: Exploring Text-free Training for Text-Driven Image Manipulation
Yueming Lyu, Tianwei Lin, Fu Li +3
Text-driven image manipulation remains challenging in training or inference flexibility. Conditional generative models depend heavily on expensive annotated training data. Meanwhil…
cs.CV2024
VCoME: Verbal Video Composition with Multimodal Editing Effects
Weibo Gong, Xiaojie Jin, Xin Li +2
Verbal videos, featuring voice-overs or text overlays, provide valuable content but present significant challenges in composition, especially when incorporating editing effects to…