27 citations · 42 across the 31 of their papers we have counts for
5 papers · 2 filters
RoboFind: Multi-Agent Personalized Object Search for People Who Are Blind or Have Low Vision
Ruiping Liu, Shaofang Quan, Qian Yin +10
Blind and low-vision users often need to locate a specific personal object rather than an arbitrary instance of the same category. The task calls for a robot that can move through…
Depth-Wise Probing and Pruning of the Planning Token in a Driving Vision-Language-Action Model
Harisankar Babu, Benjamin Coors, Christopher Lang +3
Vision-language-action (VLA) models route driving decisions through a deep language model, but it is unclear how much of that depth the action itself requires. We study a represent…
Learning to Forget -- Hierarchical Episodic Memory for Lifelong Robot Deployment
Leonard Bärmann, Joana Plewnia, Alex Waibel +1
Robots must verbalize their past experiences when users ask "Where did you put my keys?" or "Why did the task fail?" Yet maintaining life-long episodic memory (EM) from continuous…
Unified Learning of Temporal Task Structure and Action Timing for Bimanual Robot Manipulation
Christian Dreher, Patrick Dormanns, Andre Meixner +1
Bimanual manipulation requires both temporal task structure - which actions precede or overlap others - and concrete timing - when each action starts and how long it takes. Symboli…
ManipulationNet: An Infrastructure for Benchmarking Real-World Robot Manipulation with Physical Skill Challenges and Embodied Multimodal Reasoning
Yiting Chen, Kenneth Kimble, Edward H. Adelson +20
Dexterous manipulation enables robots to purposefully alter the physical world, transforming them from passive observers into active agents in unstructured environments. This capab…