Showing 2025Show all
2 papers · 1 filter
cs.RO2025
MobileUse: A GUI Agent with Hierarchical Reflection for Autonomous Mobile Operation
Ning Li, Xiangmou Qu, Jiamu Zhou +6
Recent advances in Multimodal Large Language Models (MLLMs) have enabled the development of mobile agents that can understand visual inputs and follow user instructions, unlocking…
cs.CL2025
HammerBench: Fine-Grained Function-Calling Evaluation in Real Mobile Device Scenarios
Jun Wang, Jiamu Zhou, Muning Wen +7
Evaluating the performance of LLMs in multi-turn human-agent interactions presents significant challenges, particularly due to the complexity and variability of user behavior. In t…