3 papers
cs.LG2025
Improving Multimodal Learning Balance and Sufficiency through Data Remixing
Xiaoyu Ma, Hao Chen, Yongjian Deng
Different modalities hold considerable gaps in optimization trajectories, including speeds and paths, which lead to modality laziness and modality clash when jointly training multi…
cs.CV2025
Dissecting RGB-D Learning for Improved Multi-modal Fusion
Hao Chen, Haoran Zhou, Yunshu Zhang +2
In the RGB-D vision community, extensive research has been focused on designing multi-modal learning strategies and fusion structures. However, the complementary and fusion mechani…
cs.CV2024
Prune and Repaint: Content-Aware Image Retargeting for any Ratio
Feihong Shen, Chao Li, Yifeng Geng +2
Image retargeting is the task of adjusting the aspect ratio of images to suit different display devices or presentation environments. However, existing retargeting methods often st…