agent security 1hidden state analysis 1large language models 1policy violation detection 1prompt injection 1
From the 1 of 3 linked papers with an AI index.
3 papers
cs.CR2026
PVDetector: Detecting Prompt Injection Attacks on Purpose-Specific LLM Agents through Policy-Violation Concept Analysis
Junhui Wang, Hangtao Zhang, Zhirun Zheng +5
The paper introduces PVDetector, a training‑free method that detects prompt injection attacks on purpose‑specific LLM agents by measuring alignment of hidden states with policy‑vio…
cs.AI2026
AI-Model Network: Concept, Current State and Future
Li Zhetao, Zeng Xiyu, Wang Jianhui +6
While the primary function of computers lies in computation and processing, the core value of the Internet is rooted in sharing and collaboration. Computers create the Internet, an…
cs.DC2025
Task-Agnostic Federation over Decentralized Data: Research Landscape and Visions
Wentai Wu, Ligang He, Saiqin Long +4
Increasing legislation and regulations on private and proprietary information results in scattered data sources also known as the "data islands". Although Federated Learning-based…