2 papers
cs.CV2025
MAGE: A Multi-task Architecture for Gaze Estimation with an Efficient Calibration Module
Haoming Huang, Musen Zhang, Jianxin Yang +3
Eye gaze can provide rich information on human psychological activities, and has garnered significant attention in the field of Human-Robot Interaction (HRI). However, existing gaz…
cs.RO2025
CLEA: Closed-Loop Embodied Agent for Enhancing Task Execution in Dynamic Environments
Mingcong Lei, Ge Wang, Yiming Zhao +7
Large Language Models (LLMs) exhibit remarkable capabilities in the hierarchical decomposition of complex tasks through semantic reasoning. However, their application in embodied s…