2 citations · 5 across the 9 of their papers we have counts for
9 papers
Accelerating Suffix Jailbreak attacks with Prefix-Shared KV-cache
Xinhai Wang, Shaopeng Fu, Shu Yang +3
Suffix jailbreak attacks serve as a systematic method for red-teaming Large Language Models (LLMs) but suffer from prohibitive computational costs, as a large number of candidate s…
Understanding and Mitigating Cross-lingual Privacy Leakage via Language-specific and Universal Privacy Neurons
Wenshuo Dong, Qingsong Yang, Shu Yang +5
Large Language Models (LLMs) trained on massive data capture rich information embedded in the training data. However, this also introduces the risk of privacy leakage, particularly…
Nearly Optimal Differentially Private ReLU Regression
Meng Ding, Mingxi Lei, Shaowei Wang +3
In this paper, we investigate one of the most fundamental nonconvex learning problems, ReLU regression, in the Differential Privacy (DP) model. Previous studies on private ReLU reg…
Faithful Interpretation for Graph Neural Networks
Lijie Hu, Tianhao Huang, Lu Yu +3
Currently, attention mechanisms have garnered increasing attention in Graph Neural Networks (GNNs), such as Graph Attention Networks (GATs) and Graph Transformers (GTs). It is not…
Pre-trained Encoder Inference: Revealing Upstream Encoders In Downstream Machine Learning Services
Shaopeng Fu, Xuexue Sun, Ke Qing +2
Pre-trained encoders available online have been widely adopted to build downstream machine learning (ML) services, but various attacks against these encoders also post security and…
Releasing Malevolence from Benevolence: The Menace of Benign Data on Machine Unlearning
Binhao Ma, Tianhang Zheng, Hongsheng Hu +5
Machine learning models trained on vast amounts of real or synthetic data often achieve outstanding predictive performance across various domains. However, this utility comes with…