3 papers
cs.RO2026
DUET-DINO: Simultaneous Cross-View World Modeling for Latent Planning in Robot Manipulation
Nisarga Nilavadi, Ralf Römer, Moritz Reuss +5
Action-conditioned latent world models predict future visual representations, enabling zero-shot goal-conditioned robot planning and control. However, their predictions for fine-gr…
cs.RO2026
Data and Learning Where it Matters for Contact-Rich Manipulation
Oliver Hausdörfer, Linus Schwarz, Gabor Marko +7
Learned policies trained end-to-end on large datasets often remain brittle in high-precision tasks and struggle with generalization. We find that these limitations largely stem fro…
cs.RO2026
CLARE: Continual Learning for Vision-Language-Action Models via Autonomous Adapter Routing and Expansion
Ralf Römer, Yi Zhang, Yuming Li +1
To teach robots complex manipulation tasks, a common approach is to fine-tune a pre-trained vision-language-action model (VLA) on task-specific data. However, since this recipe upd…