3 papers
cs.CV2026
Overcoming Visual Clutter in Vision Language Action Models via Concept-Gated Visual Distillation
Sangmim Song, Sarath Kodagoda, Marc Carmichael +1
Vision-Language-Action (VLA) models demonstrate impressive zero-shot generalization but frequently suffer from a "Precision-Reasoning Gap" in cluttered environments. This failure i…
cs.RO2025
Guide-LLM: An Embodied LLM Agent and Text-Based Topological Map for Robotic Guidance of People with Visual Impairments
Sangmim Song, Sarath Kodagoda, Amal Gunatilake +3
Navigation presents a significant challenge for persons with visual impairments (PVI). While traditional aids such as white canes and guide dogs are invaluable, they fall short in…
cs.MA2024
Two Heads Are Better Than One: Collaborative LLM Embodied Agents for Human-Robot Interaction
Mitchell Rosser, Marc. G Carmichael
With the recent development of natural language generation models - termed as large language models (LLMs) - a potential use case has opened up to improve the way that humans inter…