5 papers
Deform360: A Massive Multi-view Visuotactile Dataset for Deformable World Models
Hongyu Li, Wanjia Fu, Xiaoyan Cong +11
Predicting object dynamics (i.e., world modeling) is a fundamental challenge for robotic manipulation, and modeling deformable objects presents a particularly difficult case due to…
NovaPlan: Zero-Shot Long-Horizon Manipulation via Closed-Loop Video Language Planning
Jiahui Fu, Junyu Nan, Lingfeng Sun +5
Solving long-horizon tasks requires robots to integrate high-level semantic reasoning with low-level physical interaction. While vision-language models (VLMs) and video generation…
NovaFlow: Zero-Shot Manipulation via Actionable Flow from Generated Videos
Hongyu Li, Lingfeng Sun, Yafei Hu +4
Enabling robots to execute novel manipulation tasks zero-shot is a central goal in robotics. Most existing methods assume in-distribution tasks or rely on fine-tuning with embodime…
V-HOP: Visuo-Haptic 6D Object Pose Tracking
Hongyu Li, Mingxi Jia, Tuluhan Akbulut +3
Humans naturally integrate vision and haptics for robust object perception during manipulation. The loss of either modality significantly degrades performance. Inspired by this mul…
ViTa-Zero: Zero-shot Visuotactile Object 6D Pose Estimation
Hongyu Li, James Akl, Srinath Sridhar +2
Object 6D pose estimation is a critical challenge in robotics, particularly for manipulation tasks. While prior research combining visual and tactile (visuotactile) information has…