2 papers
cs.CV2026
ESPIRE: A Diagnostic Benchmark for Embodied Spatial Reasoning of Vision-Language Models
Yanpeng Zhao, Wentao Ding, Hongtao Li +2
A recent trend in vision-language models (VLMs) has been to enhance their spatial cognition for embodied domains. Despite progress, existing evaluations have been limited both in p…
cs.RO2025
In-situ Value-aligned Human-Robot Interactions with Physical Constraints
Hongtao Li, Ziyuan Jiao, Xiaofeng Liu +2
Equipped with Large Language Models (LLMs), human-centered robots are now capable of performing a wide range of tasks that were previously deemed challenging or unattainable. Howev…