1 paper · 1 filter
Qiushi Sun, Mukai Li, Zhoumianze Liu +11
Computer-using agents powered by Vision-Language Models (VLMs) have demonstrated human-like capabilities in operating digital environments like mobile platforms. While these agents…