7 papers
FingerTip 20K: A Benchmark for Proactive and Personalized Mobile LLM Agents
Qinglong Yang, Haoming Li, Haotian Zhao +4
Mobile GUI agents are becoming critical tools to improve user experience on smart devices, with multimodal large language models (MLLMs) emerging as the dominant paradigms in this…
AI Urban Scientist: Multi-Agent Collaborative Automation for Urban Research
Tong Xia, Jiankun Zhang, Ruiwen You +8
Urban research aims to understand how cities operate and evolve as complex adaptive systems. With the rapid growth of urban data and analytical methodologies, the central challenge…
AgentSwift: Efficient LLM Agent Design via Value-guided Hierarchical Search
Yu Li, Lehui Li, Zhihao Wu +5
Large language model (LLM) agents have demonstrated strong capabilities across diverse domains, yet automated agent design remains a significant challenge. Current automated agent…
AgentStealth: Reinforcing Large Language Model for Anonymizing User-generated Text
Chenyang Shao, Tianxing Li, Chenhao Pu +2
In today's digital world, casual user-generated content often contains subtle cues that may inadvertently expose sensitive personal attributes. Such risks underscore the growing im…
CrimeMind: Simulating Urban Crime with Multi-Modal LLM Agents
Qingbin Zeng, Ruotong Zhao, Jinzhu Mao +3
Modeling urban crime is an important yet challenging task that requires understanding the subtle visual, social, and cultural cues embedded in urban environments. Previous work has…
Perceive, Reflect, and Plan: Designing LLM Agent for Goal-Directed City Navigation without Instructions
Qingbin Zeng, Qinglong Yang, Shunan Dong +4
This paper considers a scenario in city navigation: an AI agent is provided with language descriptions of the goal location with respect to some well-known landmarks; By only obser…