works on

From the 2 of 28 linked papers with an AI index.

activity
20242026
collaborators

28 papers

cs.CR2026

Fingerprinting Text-to-Image Diffusion Models via Collapsed Generation

Yuanmin Huang, Chen Chen, Geng Hong +5

Proprietary text-to-image diffusion models are increasingly distributed as hosted services and downloadable checkpoints, making their intellectual property (IP) protection an incre…

cs.CR2026

FlowGuard: From Signals to Evidence for MCP Security Detection

Baichao An, Pei Chen, Geng Hong +2

The paper introduces FlowGuard, a system that detects security risks in Model Context Protocol (MCP) interactions between LLM agents and external tools by combining semantic risk a…

cs.CR2026

Rethinking MCP Security: A Large-Scale Study of Runtime MCP Servers and Security Scanner Reliability

Pei Chen, Baichao An, Mengying Wu +6

The paper introduces MCPZoo, a large collection of over 64 k Model Context Protocol (MCP) servers, and uses it to evaluate the reliability of existing security scanners for MCP ser…

cs.CV2026

AEGIS: A Mechanism-Guided Defense against Visual Synonym Jailbreaks in Text-to-Image Models

Yuanmin Huang, Zhenfei Zhang, Mi Zhang +5

Text-to-image diffusion models have achieved high visual fidelity and broad adoption, but remain vulnerable to safety violations when adversaries exploit them to synthesize illicit…

cs.CR2026

The Emergence of Autonomous Penetration Capabilities in Large Language Model-Powered AI Systems

Jiaqi Luo, Jiarun Dai, Zhile Chen +8

Nowadays, the autonomous execution of cyberattacks capable of causing substantial real-world harm is widely regarded as one of the critical red lines that frontier AI systems must…

cs.CR2026

PRISON: Unmasking the Criminal Potential of Large Language Models

Xinyi Wu, Geng Hong, Pei Chen +3

As large language models (LLMs) advance, concerns about their misconduct in complex social contexts intensify. Existing research overlooked the systematic understanding and assessm…