4 papers · 1 filter
Exploring Spatial Representation to Enhance LLM Reasoning in Aerial Vision-Language Navigation
Yunpeng Gao, Zhigang Wang, Pengfei Han +3
Aerial Vision-and-Language Navigation (VLN) is a novel task enabling Unmanned Aerial Vehicles (UAVs) to navigate in outdoor environments through natural language instructions and v…
COHERENT: Collaboration of Heterogeneous Multi-Robot System with Large Language Models
Kehui Liu, Zixin Tang, Dong Wang +3
Leveraging the powerful reasoning capabilities of large language models (LLMs), recent LLM-based robot task planning methods yield promising results. However, they mainly focus on…
AlignBot: Aligning VLM-powered Customized Task Planning with User Reminders Through Fine-Tuning for Household Robots
Zhaxizhuoma Zhaxizhuoma, Pengan Chen, Ziniu Wu +7
This paper presents AlignBot, a novel framework designed to optimize VLM-powered customized task planning for household robots by effectively aligning with user reminders. In domes…
KOI: Accelerating Online Imitation Learning via Hybrid Key-state Guidance
Jingxian Lu, Wenke Xia, Dong Wang +4
Online Imitation Learning struggles with the gap between extensive online exploration space and limited expert trajectories, hindering efficient exploration due to inaccurate rewar…