collaborators
Showing cs.CRShow all

9 papers · 1 filter

cs.CR2026

SkillShield: Prompt-Space Security Skills for LLM Coding Agents

Xiaodong Wu, Zhimin Zhao, Qi Li +4

A coding agent edits files and executes shell commands with its developer's privileges, allowing malicious requests to translate directly into harmful actions or functional malware…

cs.CR2026

EVOMAL: Self-Poisoning in Self-Evolving Coding Agents

Xiaodong Wu, Yu Shi, Qi Li +5

Self-evolving LLM coding agents write their own tools by imitating retrieved skills from shared skill libraries. We identify a vulnerability in this loop: during authoring, a retri…

cs.CR2026

MarkNull: Model-Agnostic Watermark Removal in AI-Generated Images via On-Manifold Latent Manipulation

Jie Cao, Qi Li, Zelin Zhang +4

Digital watermarking has emerged as a critical technique for provenance and copyright attribution in AI-generated imagery, yet its robustness against realistic, model-agnostic remo…

cs.CR2026

From AI-Generated Content to Agentic Action: Security and Safety Threats in Generative AI

Zelin Zhang, Qi Li, Jie Cao +2

Generative AI systems are increasingly used not only to produce content but also to retrieve data, invoke tools, and execute actions. This work examines the security and safety imp…

cs.CR2026

MarkSweep: A No-box Removal Attack on AI-Generated Image Watermarking via Noise Intensification and Frequency-aware Denoising

Jie Cao, Zelin Zhang, Qi Li +1

AI watermarking embeds invisible signals within images to provide provenance information and identify content as AI-generated. In this paper, we introduce MarkSweep, a novel waterm…

cs.CR2025

Secure and Robust Watermarking for AI-generated Images: A Comprehensive Survey

Jie Cao, Qi Li, Zelin Zhang +2

The rapid progress of Generative Artificial Intelligence (GenAI) has enabled the effortless synthesis of high-quality visual content, while simultaneously raising pressing concerns…