activity
20232026
most citedFair Text-to-Image Diffusion via Fair Mapping

2 citations · 5 across the 9 of their papers we have counts for

collaborators

9 papers

cs.CR2026

Accelerating Suffix Jailbreak attacks with Prefix-Shared KV-cache

Xinhai Wang, Shaopeng Fu, Shu Yang +3

Suffix jailbreak attacks serve as a systematic method for red-teaming Large Language Models (LLMs) but suffer from prohibitive computational costs, as a large number of candidate s…

cs.CL2025

Understanding and Mitigating Cross-lingual Privacy Leakage via Language-specific and Universal Privacy Neurons

Wenshuo Dong, Qingsong Yang, Shu Yang +5

Large Language Models (LLMs) trained on massive data capture rich information embedded in the training data. However, this also introduces the risk of privacy leakage, particularly…

cs.LG2025

Nearly Optimal Differentially Private ReLU Regression

Meng Ding, Mingxi Lei, Shaowei Wang +3

In this paper, we investigate one of the most fundamental nonconvex learning problems, ReLU regression, in the Differential Privacy (DP) model. Previous studies on private ReLU reg…

cs.LG2024

Faithful Interpretation for Graph Neural Networks

Lijie Hu, Tianhao Huang, Lu Yu +3

Currently, attention mechanisms have garnered increasing attention in Graph Neural Networks (GNNs), such as Graph Attention Networks (GATs) and Graph Transformers (GTs). It is not…

cs.LG2024

Pre-trained Encoder Inference: Revealing Upstream Encoders In Downstream Machine Learning Services

Shaopeng Fu, Xuexue Sun, Ke Qing +2

Pre-trained encoders available online have been widely adopted to build downstream machine learning (ML) services, but various attacks against these encoders also post security and…

cs.CR2024★ 2 cited

Releasing Malevolence from Benevolence: The Menace of Benign Data on Machine Unlearning

Binhao Ma, Tianhang Zheng, Hongsheng Hu +5

Machine learning models trained on vast amounts of real or synthetic data often achieve outstanding predictive performance across various domains. However, this utility comes with…