1 paper
Runze Li, Yuwen Zhai, Bo Xu +5
Contemporary GUI agents, while increasingly capable due to advances in Large Vision-Language Models (VLMs), often operate with a critical limitation: they treat each task in isolat…