2 papers
cs.CR2026
Measuring the Security of Mobile LLM Agents under Adversarial Prompts from Untrusted Third-Party Channels
Chenghao Du, Quanfeng Huang, Tingxuan Tang +3
Large Language Models (LLMs) have transformed software development, enabling AI-powered applications known as LLM-based agents that promise to automate tasks across diverse apps an…
cs.CR2026
Understanding Human-AI Collaboration in Cybersecurity Competitions
Tingxuan Tang, Nicolas Janis, Kalyn Asher Montague +6
Capture-the-Flag (CTF) competitions are increasingly becoming a testbed for evaluating AI capabilities at solving security tasks, due to the controlled environments and objective s…