5 papers
LAP: Fast LAtent Diffusion Planner for Autonomous Driving
Jinhao Zhang, Wenlong Xia, Zhexuan Zhou +3
Diffusion models have demonstrated strong capabilities for modeling human-like driving behaviors in autonomous driving, but their iterative sampling process induces substantial lat…
Hyper-DP3: Frequency-Aware Right-Sizing of 3D Diffusion Policies for Visuomotor Control
Jinhao Zhang, Zhexuan Zhou, Huizhe Li +5
Diffusion-based visuomotor policies perform well in robotic manipulation, yet current methods still inherit image-generation-style decoders and multi-step sampling. We revisit this…
Information Filtering via Variational Regularization for Robot Manipulation
Jinhao Zhang, Wenlong Xia, Yaojia Wang +6
Diffusion-based visuomotor policies built on 3D visual representations have achieved strong performance in learning complex robotic skills. However, most existing methods employ an…
Ego to World: Collaborative Spatial Reasoning in Embodied Systems via Reinforcement Learning
Heng Zhou, Li Kang, Yiran Qin +12
Understanding the world from distributed, partial viewpoints is a fundamental challenge for embodied multi-agent systems. Each agent perceives the environment through an ego-centri…
PocketDP3: Efficient Pocket-Scale 3D Visuomotor Policy
Jinhao Zhang, Zhexuan Zhou, Huizhe Li +5
Recently, 3D vision-based diffusion policies have shown strong capability in learning complex robotic manipulation skills. However, a common architectural mismatch exists in these…