1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.CR2026
When Agents Learn to Be You: Benchmarking Privacy Leakage, Impersonation Risk, and Defenses in Persona Skills
Yongli Xiang, Zhifang Zhang, Bojun Yang +4
Persona skills distill personal interaction histories into portable and executable artifacts for downstream agents. While enabling flexible personalization, this process concentrat…
cs.CR2026
Poise: Position-Aware One-Instruction Skill Injection for Silent Execution on LLM Agents
Haochang Hao, Dehai Min, Zhifang Zhang +4
Agent skills extend general-purpose agents, but their open format enables skill poisoning: a tampered skill can make an agent run an attacker's command while completing the user's…
cs.CR2022★ 1 cited
Confidence Matters: Inspecting Backdoors in Deep Neural Networks via Distribution Transfer
Tong Wang, Yuan Yao, Feng Xu +3
Backdoor attacks have been shown to be a serious security threat against deep learning models, and detecting whether a given model has been backdoored becomes a crucial task. Exist…