2 citations · 5 across the 6 of their papers we have counts for
Showing cs.CRShow all
3 papers · 1 filter
cs.CR2026
MechAudit-40: White-Box Auditing across 40 LLM Attack Mechanisms
Zhen Guo, Shanghao Shi, Shamim Yazdani +2
While LLM attacks span prompt optimization, multi-turn context manipulation, retrieval poisoning, and model backdoors, white-box defenses are typically evaluated on isolated attack…
cs.CR2026
TraceGuard: Process-Guided Firewall against Reasoning Backdoors in Large Language Models
Zhen Guo, Shanghao Shi, Hao Li +3
Large Reasoning Models (LRMs) introduce a reasoning-level attack surface: adversaries can corrupt intermediate inferences while preserving a plausible trace and an apparently benig…
cs.CR2025★ 1 cited
DarkMind: Latent Chain-of-Thought Backdoor in Customized LLMs
Zhen Guo, Shanghao Shi, Shamim Yazdani +2
With the rapid rise of personalized AI, customized large language models (LLMs) equipped with Chain of Thought (COT) reasoning now power millions of AI agents. However, their compl…