works on

From the 1 of 12 linked papers with an AI index.

activity
20242026
most citedViLAM: Distilling Vision-Language Reasoning into Attention Maps for Social Robot Navigation

1 citations · 1 across the 3 of their papers we have counts for

collaborators
Showing cs.ROShow all

9 papers · 1 filter

cs.RO2026

Paired-CSLiDAR: Height-Stratified Registration for Cross-Source Aerial-Ground LiDAR Pose Refinement

Montana Hoover, Jing Liang, Tianrui Guan +1

We introduce Paired-CSLiDAR (CSLiDAR), a cross-source aerial-ground LiDAR benchmark for single-scan pose refinement: refining a ground-scan pose within a 50 m-radius aerial crop. T…

cs.RO20261 cited

ViLAM: Distilling Vision-Language Reasoning into Attention Maps for Social Robot Navigation

Mohamed Elnoor, Kasun Weerakoon, Gershom Seneviratne +3

We introduce ViLAM, a novel method for distilling vision-language reasoning from large Vision-Language Models (VLMs) into spatial attention maps for socially compliant robot naviga…

cs.RO2025

MOSU: Autonomous Long-range Robot Navigation with Multi-modal Scene Understanding

Jing Liang, Kasun Weerakoon, Daeun Song +3

We present MOSU, a novel autonomous long-range navigation system that enhances global navigation for mobile robots through multimodal perception and on-road scene understanding. MO…

cs.RO2025

VL-TGS: Trajectory Generation and Selection using Vision Language Models in Mapless Outdoor Environments

Daeun Song, Jing Liang, Xuesu Xiao +1

We present a multi-modal trajectory generation and selection algorithm for real-world mapless outdoor navigation in human-centered environments. Such environments contain rich feat…

cs.RO2025

On the Vulnerability of LLM/VLM-Controlled Robotics

Xiyang Wu, Souradip Chakraborty, Ruiqi Xian +6

In this work, we highlight vulnerabilities in robotic systems integrating large language models (LLMs) and vision-language models (VLMs) due to input modality sensitivities. While…

cs.RO2024

VLM-Social-Nav: Socially Aware Robot Navigation through Scoring using Vision-Language Models

Daeun Song, Jing Liang, Amirreza Payandeh +3

We propose VLM-Social-Nav, a novel Vision-Language Model (VLM) based navigation approach to compute a robot's motion in human-centered environments. Our goal is to make real-time d…