1 paper
Le Wang, Zonghao Ying, Xiao Yang +7
Embodied agents powered by vision-language models (VLMs) are increasingly capable of executing complex real-world tasks, yet they remain vulnerable to hazardous instructions that m…