6 papers
Structured Evidence Selection for Weakly Supervised Video Anomaly Detection
Chenglizhao Chen, Tianxiang Nan, Wen Li +4
Weakly supervised video anomaly detection relies solely on video-level labels for training, making it difficult to accurately localize anomalous events in complex scenes. In real-w…
Multi-scale Object-Aware Gaze Estimation via Geometric Reasoning
Jiajie Mi, Xinyu Liu, Mengke Song +1
Gaze target estimation aims to predict the semantic object an observer fixates upon within an image, a task deeply rooted in the object-oriented nature of human gaze. Observers ten…
Controlling Decision Drift in Multimodal Sentiment Analysis with Missing Modalities
Chenglizhao Chen, Yuchen Cao, Xinyu Liu +3
Multimodal sentiment analysis relies on textual, acoustic, and visual signals, yet real-world data often suffer from modality missing and quality imbalance. Existing methods genera…
READ-Net: Clarifying Emotional Ambiguity via Adaptive Feature Recalibration for Audio-Visual Depression Detection
Chenglizhao Chen, Boze Li, Mengke Song +5
Depression is a severe global mental health issue that impairs daily functioning and overall quality of life. Although recent audio-visual approaches have improved automatic depres…
GenCAMO: Scene-Graph Contextual Decoupling for Environment-aware and Mask-free Camouflage Image-Dense Annotation Generation
Chenglizhao Chen, Shaojiang Yuan, Xiaoxue Lu +5
Conceal dense prediction (CDP), especially RGB-D camouflage object detection and open-vocabulary camouflage object segmentation, plays a crucial role in advancing the understanding…
Camera-based implicit mind reading by capturing higher-order semantic dynamics of human gaze within environmental context
Mengke Song, Yuge Xie, Qi Cui +5
Emotion recognition,as a step toward mind reading,seeks to infer internal states from external cues.Most existing methods rely on explicit signals-such as facial expressions,speech…