1 paper
Zhaoyan Pan, Hengyang Zhou, Xiangdong Li +5
Conversational multimodal understanding aims to infer the meaning or label of the current utterance from its preceding dialogue context together with textual, acoustic, and visual…