133 citations · 246 across the 24 of their papers we have counts for
26 papers
Rethinking Visual Embodiment Dependence in Visuomotor Policies
Hongjie Fang, Yuxuan Lu, Chenxi Wang +8
Visuomotor policies observe both the task scene and the acting embodiment, allowing embodiment-specific visual cues to influence action prediction. We study this phenomenon as visu…
HINT: Human-Intent Inception for Long-Horizon Robot Manipulation
Mingyu Mei, Haojie Xu, Shihao Jin +9
Humans can perform complex manipulations given a simple intent through an overall instruction, while continuously adapting to evolving visual observations. However, current vision-…
ChronoFlow-Policy: Unifying Past-Current-Future Interaction Flow in Visuomotor Policy Learning
Bokai Lin, Yifu Xu, Xinyu Zhan +6
Visual signals play a crucial role in policy learning by enabling models to capture object motion and interaction dynamics. Just as humans reason about actions using both past expe…
Asynchronous Multimodal Diffusion Policy Composition via Latency-Aware Guidance Fusion
Zihao He, Hongjie Fang, Shirun Tang +2
Diffusion policies have shown strong potential for robotic imitation learning, and recent extensions incorporate additional modalities to improve manipulation performance. However,…
AnyDexRT: Calibration-Free Dexterous Hand Retargeting with Few-Shot Human Guidance
Chenxi Wang, Ying Feng, Hongjie Fang +4
Teleoperation is a key interface for controlling dexterous robotic hands and collecting demonstrations for imitation learning. Its effectiveness largely depends on kinematic retarg…
X-Imitator: Spatial-Aware Imitation Learning via Bidirectional Action-Pose Interaction
Kai Xiong, Hongjie Fang, Lixin Yang +1
Effectively handling the interplay between spatial perception and action generation remains a critical bottleneck in robotic manipulation. Existing methods typically treat spatial…