1 paper · 1 filter
Jiaqi Zhang, Chen Gao, Liyuan Zhang +2
Recent advances in embodied agents with multimodal perception and reasoning capabilities based on large vision-language models (LVLMs), excel in autonomously interacting either rea…