3 papers
cs.RO2026
DREAMSTEER: Latent World Models Can Steer VLA Policies During Deployment Without Any Finetuning
Hanchen Cui, Sergio Arnaud, Arjun Majumdar +5
Pretrained vision-language-action (VLA) policies show promising zero-shot generalization, but often fail under deployment-time distribution shift, leading to decreased robustness a…
cs.RO2025
Gaussian Splatting Visual MPC for Granular Media Manipulation
Wei-Cheng Tseng, Ellina Zhang, Krishna Murthy Jatavallabhula +1
Recent advancements in learned 3D representations have enabled significant progress in solving complex robotic manipulation tasks, particularly for rigid-body objects. However, man…
cs.CV2024
PickScan: Object discovery and reconstruction from handheld interactions
Vincent van der Brugge, Marc Pollefeys, Joshua B. Tenenbaum +2
Reconstructing compositional 3D representations of scenes, where each object is represented with its own 3D model, is a highly desirable capability in robotics and augmented realit…