2 papers
cs.RO2026
Act, Sense, Act: Learning Active Perception from Large-Scale Egocentric Human Data
Jialiang Li, Yi Qiao, Yunhan Guo +2
Achieving generalizable manipulation in unconstrained environments requires the robot to proactively resolve information uncertainty, i.e., the capability of active perception. How…
cs.RO2026
IndustryNav: Exploring Spatial Reasoning of Embodied Agents in Dynamic Industrial Navigation
Yifan Li, Lichi Li, Anh Dao +15
While Visual Large Language Models (VLLMs) show great promise as embodied agents, they continue to face substantial challenges in spatial reasoning. Existing embodied benchmarks la…