3 papers
cs.CV2025
Expanding Zero-Shot Object Counting with Rich Prompts
Huilin Zhu, Senyao Li, Jingling Yuan +5
Expanding pre-trained zero-shot counting models to handle unseen categories requires more than simply adding new prompts, as this approach does not achieve the necessary alignment…
cs.AI2025
VEGAS: Towards Visually Explainable and Grounded Artificial Social Intelligence
Hao Li, Hao Fei, Zechao Hu +2
Social Intelligence Queries (Social-IQ) serve as the primary multimodal benchmark for evaluating a model's social intelligence level. While impressive multiple-choice question(MCQ)…
cs.CV2025
FocalCount: Towards Class-Count Imbalance in Class-Agnostic Counting
Huilin Zhu, Jingling Yuan, Zhengwei Yang +3
In class-agnostic object counting, the goal is to estimate the total number of object instances in an image without distinguishing between specific categories. Existing methods oft…