6 papers
Benchmarking Dynamic Affective Reasoning: A Viewer-Centric Video Emotion Dataset
Zhiyan Zhang, Peipei Song, Jinpeng Hu +3
Video emotion analysis is typically framed as a static classification problem, treating each clip as an independent labeled unit. However, such a formulation overlooks a key psycho…
StepGuard: Guarding Web Navigation via Single-Step Calibration
Zhihao Cui, Yuchen Zhang, Xiyang Sun +6
Web navigation requires agents to follow natural language goals, interact with web pages, and produce accurate answers. While recent advances leverage vision-language models and re…
To Fuse or to Drop? Dual-Path Learning for Resolving Modality Conflicts in Multimodal Emotion Recognition
Yangchen Yu, Qian Chen, Jia Li +5
Multimodal emotion recognition (MER) benefits from combining text, audio, and vision, yet standard fusion often fails when modalities conflict. Crucially, conflicts differ in resol…
Tears or Cheers? Benchmarking LLMs via Culturally Elicited Distinct Affective Responses
Chongyuan Dai, Yaling Shen, Jinpeng Hu +6
Culture serves as a fundamental determinant of human affective processing and profoundly shapes how individuals perceive and interpret emotional stimuli. Despite this intrinsic lin…
PsychEthicsBench: Evaluating Large Language Models Against Australian Mental Health Ethics
Yaling Shen, Stephanie Fong, Yiwen Jiang +9
The increasing integration of large language models (LLMs) into mental health applications necessitates robust frameworks for evaluating professional safety alignment. Current eval…
CLAIP-Emo: Parameter-Efficient Adaptation of Language-supervised models for In-the-Wild Audiovisual Emotion Recognition
Yin Chen, Jia Li, Jinpeng Hu +2
Audiovisual emotion recognition (AVER) in the wild is still hindered by pose variation, occlusion, and background noise. Prevailing methods primarily rely on large-scale domain-spe…