1 paper · 1 filter
Zeyi Sun, Ziyu Liu, Yuhang Zang +5
Repurposing large vision-language models (LVLMs) as computer use agents (CUAs) has led to substantial breakthroughs, primarily driven by human-labeled data. However, these models o…