activity
20222026
most citedConstrained Optimization with Dynamic Bound-scaling for Effective NLPBackdoor Defense

11 citations · 38 across the 14 of their papers we have counts for

collaborators

14 papers

cs.AI2026

GIF: Locally Sound Geometric Information Flow Control for LLMs

Adam Storek, Nikolaus Holzer, Zhuo Zhang +1

Large language models increasingly mediate interactions between sensitive data, untrusted inputs, and privileged actions in agentic systems, creating security and privacy risks. Th…

cs.CR2024

ASPIRER: Bypassing System Prompts With Permutation-based Backdoors in LLMs

Lu Yan, Siyuan Cheng, Xuan Chen +4

Large Language Models (LLMs) have become integral to many applications, with system prompts serving as a key mechanism to regulate model behavior and ensure ethical outputs. In thi…

cs.IR2024

Threat Behavior Textual Search by Attention Graph Isomorphism

Chanwoo Bae, Guanhong Tao, Zhuo Zhang +1

Cyber attacks cause over $1 trillion loss every year. An important task for cyber security analysts is attack forensics. It entails understanding malware behaviors and attack orig…

cs.SE2024

CodeArt: Better Code Models by Attention Regularization When Symbols Are Lacking

Zian Su, Xiangzhe Xu, Ziyang Huang +4

Transformer based code models have impressive performance in many software engineering tasks. However, their effectiveness degrades when symbols are missing or not informative. The…

cs.AI2024★ 1 cited

Rapid Optimization for Jailbreaking LLMs via Subconscious Exploitation and Echopraxia

Guangyu Shen, Siyuan Cheng, Kaiyuan Zhang +6

Large Language Models (LLMs) have become prevalent across diverse sectors, transforming human life with their extraordinary reasoning and comprehension abilities. As they find incr…

cs.CL2024

MULTIVERSE: Exposing Large Language Model Alignment Problems in Diverse Worlds

Xiaolong Jin, Zhuo Zhang, Xiangyu Zhang

Large Language Model (LLM) alignment aims to ensure that LLM outputs match with human values. Researchers have demonstrated the severity of alignment problems with a large spectrum…