12 papers
Think Proprioceptively: State-Grounded Visual Token Selection for VLA Policies
Fangyuan Wang, Peng Zhou, Jiaming Qi +4
Vision-language-action (VLA) models typically inject proprioception only as a late conditioning signal, preventing robot state from grounding instruction understanding or directing…
Open-H-Embodiment: A Large-Scale Dataset for Enabling Foundation Models in Medical Robotics
Open-H-Embodiment Consortium, :, Nigel Nelson +213
Autonomous medical robots hold promise to improve patient outcomes, reduce provider workload, democratize access to care, and enable superhuman precision. However, autonomous medic…
World Models for Robotic Manipulation: A Survey
Fangyuan Wang, Ziyuan Wang, Guorui Pei +15
Robotic manipulation depends on the ability to anticipate how actions reshape objects, contacts, and scene geometry before execution. Learned world models provide this capability b…
Failure-Aware Bimanual Teleoperation via Conservative Value Guided Assistance
Peng Zhou, Zhongxuan Li, Jinsong Wu +7
Teleoperation of high-precision manipulation is con-strained by tight success tolerances and complex contact dy-namics, which make impending failures difficult for human operators…
Non-Prehensile Tool-Object Manipulation by Integrating LLM-Based Planning and Manoeuvrability-Driven Controls
Hoi-Yin Lee, Peng Zhou, Anqing Duan +3
The ability to wield tools was once considered exclusive to human intelligence, but it's now known that many other animals, like crows, possess this capability. Yet, robotic system…
UniBiDex: A Unified Teleoperation Framework for Robotic Bimanual Dexterous Manipulation
Zhongxuan Li, Zeliang Guo, Jun Hu +4
We present UniBiDex a unified teleoperation framework for robotic bimanual dexterous manipulation that supports both VRbased and leaderfollower input modalities UniBiDex enables re…