2 citations · 2 across the 2 of their papers we have counts for
2 papers
cs.LG2026
APEX: Autonomous Policy Exploration for Self-Evolving LLM Agents
Yibo Li, Jiashuo Yang, Zhi Zheng +5
LLM agents have shown strong performance across a wide range of complex tasks, including interactive environments that require long-horizon decision making. But these agents cannot…
cs.HC2024★ 2 cited
LlamaTouch: A Faithful and Scalable Testbed for Mobile UI Task Automation
Li Zhang, Shihe Wang, Xianqing Jia +5
The emergent large language/multimodal models facilitate the evolution of mobile agents, especially in mobile UI task automation. However, existing evaluation approaches, which rel…