2 papers
cs.MM2026
State-Anchored Complete-View Distillation for Robust Conversational Multimodal Emotion Recognition
Zhaoyan Pan, Xiangdong Li, Wenke Wu +5
Conversational multimodal emotion recognition (MER) requires reliable prediction when language, acoustic, or visual observations are missing or unreliable. Many missing-modality me…
cs.MM2026
Beyond Isolated Utterances: Cue-Guided Interaction for Context-Dependent Conversational Multimodal Understanding
Zhaoyan Pan, Hengyang Zhou, Xiangdong Li +5
Conversational multimodal understanding aims to infer the meaning or label of the current utterance from its preceding dialogue context together with textual, acoustic, and visual…