16 papers
Rethink Before You Execute: Adaptive Execution for World Action Models
Feng Ye, Yiming Zhao, Yong Yu +5
World Action Models (WAMs) jointly predict future actions and the evolution of the environment. At each inference, a WAM generates a chunk of actions and the robot executes a fixed…
Interleaved POMDP Planning for Multi-Object Search in Unknown Multi-Room Household Environments
Ruochu Yang, Ziyi Xia, Huibo Zhang +6
Multi-object search in unknown household environments requires planning under extensive uncertainty - from unknown object locations to cluttered spaces with unobserved obstacles. P…
Deconfounded Lifelong Learning for Autonomous Driving via Dynamic Knowledge Spaces
Jiayuan Du, Yuebing Song, Yiming Zhao +6
End-to-End autonomous driving (E2E-AD) systems face challenges in lifelong learning, including catastrophic forgetting, difficulty in knowledge transfer across diverse scenarios, a…
TSD: A Physics-Inspired Trajectory Saliency Detector for Efficient Imitation Learning
Yiming Zhao, Gongrui Ma, Qingkai Li +1
For imitation learning in robotic manipulation, high data collection costs result in the scarcity of high quality data. In this paper, we leverage the inherent heterogeneity of tra…
ReFPO: Reflow Regularization for Flow Matching Policy Gradients
Ge Wang, Yibo Peng, Fan Feng +10
We present Reflow-regularized Flow Matching Policy Gradients (ReFPO), a simple online RL method that adds explicit Reflow regularization to FPO for efficient flow-based control. We…
MimicIK: Real-Time Generative Inverse Kinematics from Teleoperation with FK Consistency
Jiahao Yang, Shenhao Yan, Fan Feng +5
Inverse kinematics (IK) remains a critical bottleneck for real-time robot manipulation. Classical numerical solvers achieve high geometric precision but often suffer from discontin…