4 citations · 5 across the 5 of their papers we have counts for
4 papers · 1 filter
MIP-GAF: A MLLM-annotated Benchmark for Most Important Person Localization and Group Context Understanding
Surbhi Madan, Shreya Ghosh, Lownish Rai Sookha +4
Estimating the Most Important Person (MIP) in any social event setup is a challenging problem mainly due to contextual complexity and scarcity of labeled data. Moreover, the causal…
Authentic Emotion Mapping: Benchmarking Facial Expressions in Real News
Qixuan Zhang, Zhifeng Wang, Yang Liu +4
In this paper, we present a novel benchmark for Emotion Recognition using facial landmarks extracted from realistic news videos. Traditional methods relying on RGB images are resou…
A Closer Look at the Robustness of Contrastive Language-Image Pre-Training (CLIP)
Weijie Tu, Weijian Deng, Tom Gedeon
Contrastive Language-Image Pre-training (CLIP) models have demonstrated remarkable generalization capabilities across multiple challenging distribution shifts. However, there is st…
Optimizing Camera Configurations for Multi-View Pedestrian Detection
Yunzhong Hou, Xingjian Leng, Tom Gedeon +1
Jointly considering multiple camera views (multi-view) is very effective for pedestrian detection under occlusion. For such multi-view systems, it is critical to have well-designed…