activity
20242026
collaborators

7 papers

cs.CV2026

What-Meets-Where: Unified Learning of Action and Contact Localization in Images

Yuxiao Wang, Yu Lei, Wolin Liang +4

People control their bodies to establish contact with the environment. To comprehensively understand actions across diverse visual contexts, it is essential to simultaneously consi…

cs.CV2025

QueryCraft: Transformer-Guided Query Initialization for Enhanced Human-Object Interaction Detection

Yuxiao Wang, Wolin Liang, Yu Lei +3

Human-Object Interaction (HOI) detection aims to localize human-object pairs and recognize their interactions in images. Although DETR-based methods have recently emerged as the ma…

cs.CV2025

Prompt Guidance and Human Proximal Perception for HOT Prediction with Regional Joint Loss

Yuxiao Wang, Yu Lei, Zhenao Wei +4

The task of Human-Object conTact (HOT) detection involves identifying the specific areas of the human body that are touching objects. Nevertheless, current models are restricted to…

cs.CV2025

FreeA: Human-object Interaction Detection using Free Annotation Labels

Qi Liu, Yuxiao Wang, Xinyu Jiang +5

Recent human-object interaction (HOI) detection methods depend on extensively annotated image datasets, which require a significant amount of manpower. In this paper, we propose a…

cs.CV2025

A Review of Human-Object Interaction Detection

Yuxiao Wang, Yu Lei, Li Cui +3

Human-object interaction (HOI) detection plays a key role in high-level visual understanding, facilitating a deep comprehension of human activities. Specifically, HOI detection aim…

cs.CV2025

OpenVidVRD: Open-Vocabulary Video Visual Relation Detection via Prompt-Driven Semantic Space Alignment

Qi Liu, Weiying Xue, Yuxiao Wang +1

The video visual relation detection (VidVRD) task is to identify objects and their relationships in videos, which is challenging due to the dynamic content, high annotation costs,…