4 papers
PrefPoE: Advantage-Guided Preference Fusion for Learning Where to Explore
Zhihao Lin, Lin Wu, Zhen Tian +1
Exploration in reinforcement learning remains a critical challenge, as naive entropy maximization often results in high variance and inefficient policy updates. We introduce \textb…
HOI-Dyn: Learning Interaction Dynamics for Human-Object Motion Diffusion
Lin Wu, Zhixiang Chen, Jianglin Lan
Generating realistic 3D human-object interactions (HOIs) remains a challenging task due to the difficulty of modeling detailed interaction dynamics. Existing methods treat human an…
U-DiT Policy: U-shaped Diffusion Transformers for Robotic Manipulation
Linzhi Wu, Aoran Mei, Xiyue Wang +2
Diffusion-based methods have been acknowledged as a powerful paradigm for end-to-end visuomotor control in robotics. Most existing approaches adopt a Diffusion Policy in U-Net arch…
Referring to Any Person
Qing Jiang, Lin Wu, Zhaoyang Zeng +5
Humans are undoubtedly the most important participants in computer vision, and the ability to detect any individual given a natural language description, a task we define as referr…