1 paper · 1 filter
Liyun Zhang, Xuanmeng Sha, Shuqiong Wu +1
Multimodal Large Language Models (MLLMs) excel in Open-Vocabulary (OV) emotion recognition but often neglect fine-grained acoustic modeling. Existing methods typically use global a…