16 papers
Depth-Wise Probing and Pruning of the Planning Token in a Driving Vision-Language-Action Model
Harisankar Babu, Benjamin Coors, Christopher Lang +3
Vision-language-action (VLA) models route driving decisions through a deep language model, but it is unclear how much of that depth the action itself requires. We study a represent…
Contact Wasserstein Geodesics for Non-Conservative Schrödinger Bridges
Andrea Testa, Søren Hauberg, Tamim Asfour +1
The Schrödinger Bridge provides a principled framework for modeling stochastic processes between distributions; however, existing methods are limited by energy-conservation assump…
Learning to Forget -- Hierarchical Episodic Memory for Lifelong Robot Deployment
Leonard Bärmann, Joana Plewnia, Alex Waibel +1
Robots must verbalize their past experiences when users ask "Where did you put my keys?" or "Why did the task fail?" Yet maintaining life-long episodic memory (EM) from continuous…
Unified Learning of Temporal Task Structure and Action Timing for Bimanual Robot Manipulation
Christian Dreher, Patrick Dormanns, Andre Meixner +1
Temporal task structure is fundamental for bimanual manipulation: a robot must not only know that one action precedes or overlaps another, but also when each action should occur an…
Taxonomy-aware Dynamic Motion Generation on Hyperbolic Manifolds
Luis Augenstein, Noémie Jaquier, Tamim Asfour +1
Human-like motion generation for robots often draws inspiration from biomechanical studies, which often categorize complex human motions into hierarchical taxonomies. While these t…
Diffusion-Based Impedance Learning for Contact-Rich Manipulation Tasks
Noah Geiger, Tamim Asfour, Neville Hogan +1
Learning-based methods excel at robot motion generation but remain limited in contact-rich physical interaction. Impedance control provides stable and safe contact behavior but req…