1 paper · 1 filter
Ning Li, Xiangmou Qu, Jiamu Zhou +6
Recent advances in Multimodal Large Language Models (MLLMs) have enabled the development of mobile agents that can understand visual inputs and follow user instructions, unlocking…