11 papers
G0.5: One Autoregressive Stream for Robot Reasoning and Action
Yicheng Liu, Zibin Dong, Baijun Ye +24
The prevailing recipe for Vision-Language-Action (VLA) models couples a pretrained VLM with a separately trained flow-matching action expert. This makes the VLM a context encoder r…
OMG: Omni-Modal Motion Generation for Generalist Humanoid Control
Siqiao Huang, Kun-Ying Lee, Dongming Qiao +5
Humanoid whole-body control has made significant progress in recent years, yet existing approaches remain limited to few-skill policies with heavy reward engineering, or motion tra…
Generative Control as Optimization: Time Unconditional Flow Matching for Adaptive and Robust Robotic Control
Zunzhe Zhang, Runhan Huang, Yicheng Liu +3
Diffusion models and flow matching have become a cornerstone of robotic imitation learning, yet they suffer from a structural inefficiency where inference is often bound to a fixed…
TTT-Parkour: Rapid Test-Time Training for Perceptive Robot Parkour
Shaoting Zhu, Baijun Ye, Jiaxuan Wang +5
Achieving highly dynamic humanoid parkour on unseen, complex terrains remains a challenge in robotics. Although general locomotion policies demonstrate capabilities across broad te…
Hiking in the Wild: A Scalable Perceptive Parkour Framework for Humanoids
Shaoting Zhu, Ziwen Zhuang, Mengjie Zhao +2
Achieving robust humanoid hiking in complex, unstructured environments requires transitioning from reactive proprioception to proactive perception. However, integrating exterocepti…
Deep Whole-body Parkour
Ziwen Zhuang, Shaoting Zhu, Mengjie Zhao +1
Current approaches to humanoid control generally fall into two paradigms: perceptive locomotion, which handles terrain well but is limited to pedal gaits, and general motion tracki…