2 papers
cs.CV2026
AlloEgo-VLM: Disambiguating Allocentric and Egocentric Reference Frames in Vision-Language Models
Kuan-Lin Chen, Tzu-Ti Wei, Chao-Chi Liao +2
This study investigates the challenge of ambiguity faced by Vision-Language Models (VLMs) in understanding spatial semantics. Spatial cognition, shaped by cognitive psychology, spa…
cs.RO2024
Resolving Positional Ambiguity in Dialogues by Vision-Language Models for Robot Navigation
Kuan-Lin Chen, Tzu-Ti Wei, Li-Tzu Yeh +3
We consider an autonomous navigation robot that can accept human commands through natural language to provide services in an indoor environment. These natural language commands may…