3 papers
cs.CR2026
PVDetector: Detecting Prompt Injection Attacks on Purpose-Specific LLM Agents through Policy-Violation Concept Analysis
Junhui Wang, Hangtao Zhang, Zhirun Zheng +5
Large language models (LLMs) are increasingly deployed as purpose-specific agents to handle domain-specific tasks such as customer service and code generation. These agents are exp…
cs.SE2026
The Hitchhiker's Guide to Program Analysis, Part III: Mostly Harmless LLMs
Haonan Li, Tianyang Zhou, Manu Sridharan +2
LLMs are increasingly used in bug analysis to reason about code and judge whether a potential bug can be triggered in realistic execution contexts, with recent work showing promisi…
cs.CR2026
Defending Jailbreak Attacks on Large Language Models via Manifold Trajectory Kinetics
Hangtao Zhang, Yucheng Zhao, Sishun Liu +8
Jailbreak prompts can bypass alignment guardrails in large language models (LLMs) and elicit unsafe outputs, making reliable deployment-time detection critical. Prior detection app…