1 paper · 1 filter
Chiyu Chen, Xinhao Song, Yunkai Chai +7
Vision-Language Models (VLMs) are increasingly deployed as autonomous agents to navigate mobile graphical user interfaces (GUIs). Operating in dynamic on-device ecosystems, which i…