Showing cs.CRShow all
3 papers · 1 filter
cs.CR2026
AgentLeak: Cloning Stronger LLM Agent Capabilities onto Weaker Agents Beyond Skill Stealing
Xiaoting Lyu, Yuhong Wu, Yufei Han +5
Large language model (LLM) agents increasingly achieve long-horizon tasks by combining foundation models with explicit skills and implicit procedural knowledge acquired through exe…
cs.CR2026
AgentTrap: Measuring Runtime Trust Failures in Third-Party Agent Skills
Haomin Zhuang, Hanwen Xing, Yujun Zhou +5
Third-party skills are becoming the package ecosystem for LLM agents. They package natural-language instructions, helper scripts, templates, documents, and service configuration in…
cs.CR2025
Persistent Backdoor Attacks under Continual Fine-Tuning of LLMs
Jing Cui, Yufei Han, Jianbin Jiao +1
Backdoor attacks embed malicious behaviors into Large Language Models (LLMs), enabling adversaries to trigger harmful outputs or bypass safety controls. However, the persistence of…