2 citations · 2 across the 6 of their papers we have counts for
10 papers
Gripper-aware Vision Language Action Models
Hanyi Zhang, Zihong Luo, Tianyu Li +16
Vision language action models (VLAs) have advanced general purpose robotic grasping and manipulation by enabling robots to interpret visual observations and natural language instru…
Rheos: Modelling Continuous Motion Dynamics in Hierarchical 3D Scene Graphs
Iacopo Catalano, Francesco Verdoja, Javier Civera +2
3D Scene Graphs (3DSGs) provide hierarchical, multi-resolution abstractions that encode the geometric and semantic structure of an environment, yet their treatment of dynamics rema…
EgoMoD: Predicting Global Maps of Dynamics from Local Egocentric Observations
Iacopo Catalano, David Morilla-Cabello, Jorge Pena-Queralta +1
Efficient navigation in dynamic environments requires anticipating how motion patterns evolve beyond the robot's immediate perceptual range, enabling preemptive rather than purely…
Aion: Towards Hierarchical 4D Scene Graphs with Temporal Flow Dynamics
Iacopo Catalano, Eduardo Montijano, Javier Civera +2
Autonomous navigation in dynamic environments requires spatial representations that capture both semantic structure and temporal evolution. 3D Scene Graphs (3DSGs) provide hierarch…
Towards Embodied Agentic AI: Review and Classification of LLM- and VLM-Driven Robot Autonomy and Interaction
Sahar Salimpour, Lei Fu, Kajetan RachwaÅ +8
Foundation models, including large language models (LLMs) and vision-language models (VLMs), have recently enabled novel approaches to robot autonomy and human-robot interfaces. In…
Follow-Me in Micro-Mobility with End-to-End Imitation Learning
Sahar Salimpour, Iacopo Catalano, Tomi Westerlund +2
Autonomous micro-mobility platforms face challenges from the perspective of the typical deployment environment: large indoor spaces or urban areas that are potentially crowded and…