1 paper
Ziyi Bai, Siqi Li, Tinglei Huang +1
Recent studies have shown that multimodal large language models (MLLMs) can serve as embodied agents, translating language instructions and visual observations into executable plan…