2 papers
cs.CV2025
DisentangleFormer: Spatial-Channel Decoupling for Multi-Channel Vision
Jiashu Liao, Pietro Liò, Marc de Kamps +1
Vision Transformers face a fundamental limitation: standard self-attention jointly processes spatial and channel dimensions, leading to entangled representations that prevent indep…
cs.MM2025
Sync-TVA: A Graph-Attention Framework for Multimodal Emotion Recognition with Cross-Modal Fusion
Zeyu Deng, Yanhui Lu, Jiashu Liao +2
Multimodal emotion recognition (MER) is crucial for enabling emotionally intelligent systems that perceive and respond to human emotions. However, existing methods suffer from limi…