3 papers
cs.CV2025
EmotionHallucer: Evaluating Emotion Hallucinations in Multimodal Large Language Models
Bohao Xing, Xin Liu, Guoying Zhao +3
Emotion understanding is a critical yet challenging task. Recent advances in Multimodal Large Language Models (MLLMs) have significantly enhanced their capabilities in this area. H…
cs.CV2025
FSBench: A Figure Skating Benchmark for Advancing Artistic Sports Understanding
Rong Gao, Xin Liu, Zhuozhao Hu +4
Figure skating, known as the "Art on Ice," is among the most artistic sports, challenging to understand due to its blend of technical elements (like jumps and spins) and overall ar…
cs.CV2025
AU-TTT: Vision Test-Time Training model for Facial Action Unit Detection
Bohao Xing, Kaishen Yuan, Zitong Yu +2
Facial Action Units (AUs) detection is a cornerstone of objective facial expression analysis and a critical focus in affective computing. Despite its importance, AU detection faces…