1 paper · 1 filter
Bingrui Sima, Lizhong Wang, Xiaoya Lu +2
While Vision-Language Models (VLMs) have empowered embodied agents to execute complex household tasks, they struggle to proactively handle dynamically emerging hazards during close…