4 papers · 1 filter
CoAdapt-GUI: Joint Workflow Context and Policy Adaptation for Unseen GUI Applications
Linqiang Guo, Li Gu, Zihuan Jiang +8
Mobile GUI agents remain brittle when deployed to applications absent from source training. We study novel-app generalization under a limited target interaction budget and without…
Benchmarking LLM Judges for Mobile Agent Evaluation
Ziqiang Wang, Ziqiang Wan, Li Gu +5
Mobile agent benchmarks increasingly rely on LLM-based judges to evaluate task completion, yet the reliability of these judges on mobile agent trajectories remains largely unexamin…
StepReflect: Structured UI Transition Reflection for Mobile GUI Agents
Linqiang Guo, Wei Liu, Li Gu +3
Autonomous mobile GUI agents require accurate action reflection for reliable long-horizon execution. Existing approaches rely on open-ended multimodal reasoning after each action,…
ETR: Entropy Trend Reward for Efficient Chain-of-Thought Reasoning
Xuan Xiong, Huan Liu, Li Gu +4
Chain-of-thought (CoT) reasoning improves large language model performance on complex tasks, but often produces excessively long and inefficient reasoning traces. Existing methods…