4 citations · 8 across the 9 of their papers we have counts for
Showing 2025Show all
2 papers · 1 filter
cs.CV2025
ViDiC: Video Difference Captioning
Jiangtao Wu, Shihao Li, Zhaozhou Bian +7
Understanding visual differences between dynamic scenes requires the comparative perception of compositional, spatial, and temporal changes--a capability that remains underexplored…
cs.CV2025
Milmer: a Framework for Multiple Instance Learning based Multimodal Emotion Recognition
Zaitian Wang, Jian He, Yu Liang +8
Emotions play a crucial role in human behavior and decision-making, making emotion recognition a key area of interest in human-computer interaction (HCI). This study addresses the…