6 papers · 1 filter
10 Open Challenges Steering the Future of Vision-Language-Action Models
Soujanya Poria, Navonil Majumder, Chia-Yu Hung +7
Due to their ability of follow natural language instructions, vision-language-action (VLA) models are increasingly prevalent in the embodied AI arena, following the widespread succ…
Open Scene Graphs for Open-World Object-Goal Navigation
Joel Loo, Zhanxin Wu, David Hsu
How can we build general-purpose robot systems for open-world semantic navigation, e.g., searching a novel environment for a target object specified in natural language? To tackle…
Neural Randomized Planning for Whole Body Robot Motion
Yunfan Lu, Yuchen Ma, David Hsu +1
Robot motion planning has made vast advances over the past decades, but the challenge remains: robot mobile manipulators struggle to plan long-range whole-body motion in common hou…
IntentionNet: Map-Lite Visual Navigation at the Kilometre Scale
Wei Gao, Bo Ai, Joel Loo +2
This work explores the challenges of creating a scalable and robust robot navigation system that can traverse both indoor and outdoor environments to reach distant goals. We propos…
Open Scene Graphs for Open World Object-Goal Navigation
Joel Loo, Zhanxin Wu, David Hsu
How can we build robots for open-world semantic navigation tasks, like searching for target objects in novel scenes? While foundation models have the rich knowledge and generalisat…
Scene Action Maps: Behavioural Maps for Navigation without Metric Information
Joel Loo, David Hsu
Humans are remarkable in their ability to navigate without metric information. We can read abstract 2D maps, such as floor-plans or hand-drawn sketches, and use them to navigate in…