3 papers
cs.CR2026
GUIGuard-Bench: Toward a General Evaluation for Privacy-Preserving GUI Agents
Yanxi Wang, Zhiling Zhang, Wenbo Zhou +6
As GUI agents increasingly rely on screenshots to perceive and operate digital environments, they may inadvertently expose sensitive information such as identities, accounts, locat…
cs.AI2025
GTM: Simulating the World of Tools for AI Agents
Zhenzhen Ren, Xinpeng Zhang, Zhenxing Qian +4
The integration of external tools is pivotal for empowering Large Language Model (LLM) agents with real-world capabilities. However, training these agents through direct, continuou…
cs.HC2025
Is Passive Expertise-Based Personalization Enough? A Case Study in AI-Assisted Test-Taking
Li Siyan, Jason Zhang, Akash Maharaj +2
Novice and expert users have different systematic preferences in task-oriented dialogues. However, whether catering to these preferences actually improves user experience and task…