Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
Zero-MELO: Test-Time Evidence Calibration with Multimodal LLMs for Zero-Shot Micro-Gesture Recognition
Chengyan Wang, Hanliang Xie, Yueyi Yang +1
While Multimodal Large Language Models (MLLMs) excel in general video understanding, their capability in fine-grained and motion-centric tasks remains limited. This limitation is p…
cs.CV2026
iMiGUE-3K: A Large-Scale Benchmark for Micro-Gesture Analysis with Self-Supervised Learning
Chengyan Wang, Haoyu Chen, Hui Wei +3
Emotion understanding is a fundamental challenge in affective computing and artificial intelligence. While existing approaches predominantly focus on facial expressions and speech,…