4 papers · 1 filter
EdgeMem: LLM-Free Agent Memory Construction and Retrieval via Evidence-Preserving Multi-Anchor Hypergraph
Zeyang Cui, Jiannong Cao, Zhiyuan Wen +3
Agent memory allows LLM agents to use earlier interactions when answering new queries. Existing methods often compress interaction histories into summaries or other LLM-generated r…
Automated Trajectory Evaluation for Mobile Agents via Step-Level Consequence Reasoning and Aggregation
Pengshuai Yang, Zijing Gao, Xue Yu +3
Evaluating language-guided mobile agents has recently shifted from rule-based to model-based approaches to achieve scalable and automated assessments. However, existing holistic ev…
SeerGuard: A Safety Framework for Mobile GUI Agents via World Model Prediction
Xue Yu, Bo Yuan, Kailin Zhao +3
Mobile graphical user interface (GUI) agents have demonstrated remarkable capabilities in automating complex tasks, yet they introduce critical safety risks because a single errone…
RMA: an Agentic System for Research-Level Mathematical Problems
Zelin Zhao, Bo Yuan, Jaemoo Choi +1
We present , an agentic framework for automated reasoning on research-level mathematical problems. Unlike prior studies centered on competition…