3 papers
cs.CV2026
Diverse via bounded Agreement: Geometric Regularization for Multimodal Fusion
Zixuan Xia, Hao Wang, Pengcheng Weng +4
Multimodal fusion is often treated as an optimization-balancing problem, where training signals are adjusted to prevent one modality from dominating the others. However, balanced o…
cs.CV2025
HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training
Xuecheng Wu, Danlei Huang, Heli Sun +9
Advances in Generative AI have made video-level deepfake detection increasingly challenging, exposing the limitations of current detection techniques. In this paper, we present HOL…
cs.CV2024
Gradient-Guided Modality Decoupling for Missing-Modality Robustness
Hao Wang, Shengda Luo, Guosheng Hu +1
Multimodal learning with incomplete input data (missing modality) is practical and challenging. In this work, we conduct an in-depth analysis of this challenge and find that modali…