1 paper · 1 filter
Yanda Li, Chi Zhang, Wenjia Jiang +6
With the advancement of Multimodal Large Language Models (MLLM), LLM-driven visual agents are increasingly impacting software interfaces, particularly those with graphical user int…