activity
20242026
collaborators

7 papers

cs.HC2026

FingerTip 20K: A Benchmark for Proactive and Personalized Mobile LLM Agents

Qinglong Yang, Haoming Li, Haotian Zhao +4

Mobile GUI agents are becoming critical tools to improve user experience on smart devices, with multimodal large language models (MLLMs) emerging as the dominant paradigms in this…

cs.CY2025

AI Urban Scientist: Multi-Agent Collaborative Automation for Urban Research

Tong Xia, Jiankun Zhang, Ruiwen You +8

Urban research aims to understand how cities operate and evolve as complex adaptive systems. With the rapid growth of urban data and analytical methodologies, the central challenge…

cs.CL2025

AgentSwift: Efficient LLM Agent Design via Value-guided Hierarchical Search

Yu Li, Lehui Li, Zhihao Wu +5

Large language model (LLM) agents have demonstrated strong capabilities across diverse domains, yet automated agent design remains a significant challenge. Current automated agent…

cs.CL2025

AgentStealth: Reinforcing Large Language Model for Anonymizing User-generated Text

Chenyang Shao, Tianxing Li, Chenhao Pu +2

In today's digital world, casual user-generated content often contains subtle cues that may inadvertently expose sensitive personal attributes. Such risks underscore the growing im…

cs.AI2025

CrimeMind: Simulating Urban Crime with Multi-Modal LLM Agents

Qingbin Zeng, Ruotong Zhao, Jinzhu Mao +3

Modeling urban crime is an important yet challenging task that requires understanding the subtle visual, social, and cultural cues embedded in urban environments. Previous work has…

cs.AI2024

Perceive, Reflect, and Plan: Designing LLM Agent for Goal-Directed City Navigation without Instructions

Qingbin Zeng, Qinglong Yang, Shunan Dong +4

This paper considers a scenario in city navigation: an AI agent is provided with language descriptions of the goal location with respect to some well-known landmarks; By only obser…