From the 1 of 13 linked papers with an AI index.
13 papers
ARGUS: Aligning Robot Scene Geometry Under Shifting Views with Large 3D Vision Models
Rishik Sathua, Haonan Chen, Katherine Driggs-Campbell
Large-scale visuomotor policies have demonstrated impressive performance across a wide range of robot manipulation tasks. However, despite this success, manipulation polices often…
Ordered Action Tokens for Visuomotor Policy Learning
Chaoqi Liu, Yue Zhao, Haonan Chen +4
Action tokenization maps continuous robot action chunks to discrete tokens and has become an important interface for modern visuomotor policies. Existing approaches either rely on…
Masked Visual Actions for Unified World Modeling
Hadi Alzayer, Wenlong Huang, Haonan Chen +8
Video models absorb rich priors over how the visual world moves, interacts, and responds to contact, making them promising substrates for robotic world modeling. The central challe…
DriftWorld: Fast World Modeling through Drifting
Susie Lu, Haonan Chen, Weirui Ye +1
The paper introduces DriftWorld, an action‑conditioned world model that uses a drifting generative approach to produce future frames in a single forward pass, enabling fast (30+ fp…
B-spline Policy: Accelerating Manipulation Policies via B-spline Action Representations
Xiaoshen Han, Haoyu Xiong, Haonan Chen +4
In this work, we present B-spline Policy (BSP), an action representation designed for accelerating robot manipulation policies. Rather than predicting discrete-time action chunks,…
You Only Touch Once: 6-DoF Object Pose Estimation from Single Tactile Contact
Pengfei Ye, Yuxiang Ma, Haonan Chen +5
Accurate 6-DoF object pose estimation is fundamental to robotic manipulation, yet vision-based methods often fail under occlusion, poor lighting, and reflective or transparent surf…