collaborators
Showing cs.ROShow all

14 papers · 1 filter

cs.RO2026

D-VLC: Decentralized Vision-Language Collaboration for Heterogeneous Embodied Multi-Robot Systems in Unknown Environments

Yuan Zhou, Ruitong Lin, Shen Wang +6

Multi-robot systems, particularly heterogeneous robot swarms, can improve the efficiency of complex task execution through parallel collaboration and complementary capabilities. Ho…

cs.RO2026

PathPainter: Transferring the Generalization Ability of Image Generation Models to Embodied Navigation

Yijin Wang, Yuru Tian, Xijie Huang +5

Bird's-eye-view (BEV) images have been widely demonstrated to provide valuable prior information for navigation. Given the global information provided by such views, two key challe…

cs.RO2026

Towards Precise Intent-Aligned VLA Aerial Navigation via Expert-Guided GRPO

Tianyang Chen, Wenjun Li, Xin zhou +2

Vision-Language-Action (VLA) models offer a promising end-to-end paradigm for unmanned aerial vehicles (UAVs) to accomplish complex tasks specified by fine-grained instructions. Ho…

cs.RO2026

FlyMirage: A Fully Automated Generation Pipeline for Diverse and Scalable UAV Flight Data via Generative World Model

Jinhan Li, Xijie Huang, Zhaoqi Wang +7

In the field of Vision-Language Navigation (VLN), aerial datasets remain limited in their ability to combine scale, diversity, and realism, often relying on either costly real-worl…

cs.RO2026

Precise Aggressive Aerial Maneuvers with Sensorimotor Policies

Tianyue Wu, Guangtong Xu, Zihan Wang +6

Precise aggressive maneuvers with lightweight onboard sensors remains a key bottleneck in fully exploiting the maneuverability of drones. Such maneuvers are critical for expanding…

cs.RO2026

NavDreamer: Video Models as Zero-Shot 3D Navigators

Xijie Huang, Weiqi Gai, Tianyue Wu +5

Previous Vision-Language-Action models face critical limitations in navigation: scarce, diverse data from labor-intensive collection and static representations that fail to capture…