3 papers
cs.CV2025
Collaborating Vision, Depth, and Thermal Signals for Multi-Modal Tracking: Dataset and Algorithm
Xue-Feng Zhu, Tianyang Xu, Yifan Pan +5
Existing multi-modal object tracking approaches primarily focus on dual-modal paradigms, such as RGB-Depth or RGB-Thermal, yet remain challenged in complex scenarios due to limited…
cs.CV2025
OCCO: LVM-guided Infrared and Visible Image Fusion Framework based on Object-aware and Contextual COntrastive Learning
Hui Li, Congcong Bian, Zeyang Zhang +3
Image fusion is a crucial technique in the field of computer vision, and its goal is to generate high-quality fused images and improve the performance of downstream tasks. However,…
cs.CV2023
TextFusion: Unveiling the Power of Textual Semantics for Controllable Image Fusion
Chunyang Cheng, Tianyang Xu, Xiao-Jun Wu +4
Advanced image fusion methods are devoted to generating the fusion results by aggregating the complementary information conveyed by the source images. However, the difference in th…