papers

Publications (21)

cs.CV2022

Phrase-Based Affordance Detection via Cyclic Bilateral Interaction

Liangsheng Lu, Wei Zhai, Hongchen Luo +2

Affordance detection, which refers to perceiving objects with potential action possibilities in images, is a challenging task since the possible affordance depends on the person's…

cs.LG2026

IIB-LPO: Latent Policy Optimization via Iterative Information Bottleneck

Huilin Deng, Hongchen Luo, Yue Zhu +8

Recent advances in Reinforcement Learning with Verifiable Rewards (RLVR) for Large Language Model (LLM) reasoning have been hindered by a persistent challenge: exploration collapse…

cs.CV2024

Intention-driven Ego-to-Exo Video Generation

Hongchen Luo, Kai Zhu, Wei Zhai +1

Ego-to-exo video generation refers to generating the corresponding exocentric video according to the egocentric video, providing valuable applications in AR/VR and embodied AI. Ben…

cs.CV2024

PEAR: Phrase-Based Hand-Object Interaction Anticipation

Zichen Zhang, Hongchen Luo, Wei Zhai +2

First-person hand-object interaction anticipation aims to predict the interaction process over a forthcoming period based on current scenes and prompts. This capability is crucial…

cs.CV2026

VMAD: Visual-enhanced Multimodal Large Language Model for Zero-Shot Anomaly Detection

Huilin Deng, Hongchen Luo, Wei Zhai +2

Zero-shot anomaly detection (ZSAD) recognizes and localizes anomalies in previously unseen objects by establishing feature mapping between textual prompts and inspection images, de…

cs.CV2021

Learning Visual Affordance Grounding from Demonstration Videos

Hongchen Luo, Wei Zhai, Jing Zhang +2

Visual affordance grounding aims to segment all possible interaction regions between people and objects from an image/video, which is beneficial for many applications, such as robo…