1 paper
Chiyu Chen, Xinhao Song, Yunkai Chai +7
Vision-Language Models (VLMs) are increasingly deployed as autonomous agents to navigate mobile graphical user interfaces (GUIs). Operating in dynamic on-device ecosystems, which i…