4 papers
State-Anchored Complete-View Distillation for Robust Conversational Multimodal Emotion Recognition
Zhaoyan Pan, Xiangdong Li, Wenke Wu +5
Conversational multimodal emotion recognition (MER) requires reliable prediction when language, acoustic, or visual observations are missing or unreliable. Many missing-modality me…
Beyond Isolated Utterances: Cue-Guided Interaction for Context-Dependent Conversational Multimodal Understanding
Zhaoyan Pan, Hengyang Zhou, Xiangdong Li +5
Conversational multimodal understanding aims to infer the meaning or label of the current utterance from its preceding dialogue context together with textual, acoustic, and visual…
Euler-inspired Decoupling Neural Operator for Efficient Pansharpening
Anqi Zhu, Mengting Ma, Yizhen Jiang +4
Pansharpening aims to synthesize high-resolution multispectral (HR-MS) images by fusing the spatial textures of panchromatic (PAN) images with the spectral information of low-resol…
MAUGen: A Unified Diffusion Approach for Multi-Identity Facial Expression and AU Label Generation
Xiangdong Li, Ye Lou, Ao Gao +2
The lack of large-scale, demographically diverse face images with precise Action Unit (AU) occurrence and intensity annotations has long been recognized as a fundamental bottleneck…