1 paper · 1 filter
Zichen Zhu, Hao Tang, Yansi Li +13
Existing Multimodal Large Language Model (MLLM)-based agents face significant challenges in handling complex GUI (Graphical User Interface) interactions on devices. These challenge…