1 paper
Xuan Wang, Siyuan Liang, Zhe Liu +5
Mobile agents powered by vision-language models (VLMs) are increasingly adopted for tasks such as UI automation and camera-based assistance. These agents are typically fine-tuned u…