From the 2 of 28 linked papers with an AI index.
28 papers
Fingerprinting Text-to-Image Diffusion Models via Collapsed Generation
Yuanmin Huang, Chen Chen, Geng Hong +5
Proprietary text-to-image diffusion models are increasingly distributed as hosted services and downloadable checkpoints, making their intellectual property (IP) protection an incre…
FlowGuard: From Signals to Evidence for MCP Security Detection
Baichao An, Pei Chen, Geng Hong +2
The paper introduces FlowGuard, a system that detects security risks in Model Context Protocol (MCP) interactions between LLM agents and external tools by combining semantic risk a…
Rethinking MCP Security: A Large-Scale Study of Runtime MCP Servers and Security Scanner Reliability
Pei Chen, Baichao An, Mengying Wu +6
The paper introduces MCPZoo, a large collection of over 64 k Model Context Protocol (MCP) servers, and uses it to evaluate the reliability of existing security scanners for MCP ser…
AEGIS: A Mechanism-Guided Defense against Visual Synonym Jailbreaks in Text-to-Image Models
Yuanmin Huang, Zhenfei Zhang, Mi Zhang +5
Text-to-image diffusion models have achieved high visual fidelity and broad adoption, but remain vulnerable to safety violations when adversaries exploit them to synthesize illicit…
The Emergence of Autonomous Penetration Capabilities in Large Language Model-Powered AI Systems
Jiaqi Luo, Jiarun Dai, Zhile Chen +8
Nowadays, the autonomous execution of cyberattacks capable of causing substantial real-world harm is widely regarded as one of the critical red lines that frontier AI systems must…
PRISON: Unmasking the Criminal Potential of Large Language Models
Xinyi Wu, Geng Hong, Pei Chen +3
As large language models (LLMs) advance, concerns about their misconduct in complex social contexts intensify. Existing research overlooked the systematic understanding and assessm…