2 papers
cs.CV2026
Entity-Guided Multi-Task Learning for Infrared and Visible Image Fusion
Wenyu Shao, Hongbo Liu, Yunchuan Ma +1
Existing text-driven infrared and visible image fusion approaches often rely on textual information at the sentence level, which can lead to semantic noise from redundant text and…
cs.LG2025
Simplifying CLIP: Unleashing the Power of Large-Scale Models on Consumer-level Computers
Hongbo Liu
Contrastive Language-Image Pre-training (CLIP) has attracted a surge of attention for its superior zero-shot performance and excellent transferability to downstream tasks. However,…