1 citations · 1 across the 2 of their papers we have counts for
3 papers
cs.CV2026
Overcoming Visual Clutter in Vision Language Action Models via Concept-Gated Visual Distillation
Sangmim Song, Sarath Kodagoda, Marc Carmichael +1
Vision-Language-Action (VLA) models demonstrate impressive zero-shot generalization but frequently suffer from a "Precision-Reasoning Gap" in cluttered environments. This failure i…
cs.MA2024★ 1 cited
Two Heads Are Better Than One: Collaborative LLM Embodied Agents for Human-Robot Interaction
Mitchell Rosser, Marc. G Carmichael
With the recent development of natural language generation models - termed as large language models (LLMs) - a potential use case has opened up to improve the way that humans inter…
cs.RO2024
Guide-LLM: An Embodied LLM Agent and Text-Based Topological Map for Robotic Guidance of People with Visual Impairments
Sangmim Song, Sarath Kodagoda, Amal Gunatilake +3
Navigation presents a significant challenge for persons with visual impairments (PVI). While traditional aids such as white canes and guide dogs are invaluable, they fall short in…