3 papers
cs.HC2026
MobileForge: Annotation-Free Adaptation for Mobile GUI Agents with Hierarchical Feedback-Guided Policy Optimization
Guangyi Liu, Pengxiang Zhao, Gao Wu +9
MLLM-based mobile GUI agents have made substantial progress in UI understanding and action execution, but adapting them to real target apps remains costly because mobile apps are n…
cs.HC2025
TaskSense: Cognitive Chain Modeling and Difficulty Estimation for GUI Tasks
Yiwen Yin, Zhian Hu, Xiaoxi Xu +4
Measuring GUI task difficulty is crucial for user behavior analysis and agent capability evaluation. Yet, existing benchmarks typically quantify difficulty based on motor actions (…
cs.HC2025
LLMartini: Seamless and Interactive Leveraging of Multiple LLMs through Comparison and Composition
Yingtian Shi, Jinda Yang, Yuhan Wang +4
The growing diversity of large language models (LLMs) means users often need to compare and combine outputs from different models to obtain higher-quality or more comprehensive res…