agent security 1hidden state analysis 1large language models 1policy violation detection 1prompt injection 1
From the 1 of 2 linked papers with an AI index.
2 papers
cs.CR2026
PVDetector: Detecting Prompt Injection Attacks on Purpose-Specific LLM Agents through Policy-Violation Concept Analysis
Junhui Wang, Hangtao Zhang, Zhirun Zheng +5
The paper introduces PVDetector, a training‑free method that detects prompt injection attacks on purpose‑specific LLM agents by measuring alignment of hidden states with policy‑vio…
cond-mat.str-el2026
The ground ytterbium doublet in h-YbMnO3 and the related low-temperature peculiarities of the compound
S. A. Klimin, N. D. Molchanova, N. N. Kuzmin +3
We have performed detailed temperature-dependent study of optical f-f transitions of the Yb3+ ions in h-YbMnO3 by means of Fourier-transform spectroscopy. The splitting of the grou…