Showing cs.AIShow all
2 papers · 1 filter
cs.AI2025
Does Chain-of-Thought Reasoning Help Mobile GUI Agent? An Empirical Study
Li Zhang, Longxi Gao, Mengwei Xu
Reasoning capabilities have significantly improved the performance of vision-language models (VLMs) in domains such as mathematical problem-solving, coding, and visual question-ans…
cs.AI2024
DroidCall: A Dataset for LLM-powered Android Intent Invocation
Weikai Xie, Li Zhang, Shihe Wang +2
The growing capabilities of large language models in natural language understanding significantly strengthen existing agentic systems. To power performant on-device mobile agents f…