20 citations · 35 across the 3 of their papers we have counts for
3 papers
cs.CR2025
"To Survive, I Must Defect": Jailbreaking LLMs via the Game-Theory Scenarios
Zhen Sun, Zongmin Zhang, Deqi Liang +8
As LLMs become more common, non-expert users can pose risks, prompting extensive research into jailbreak attacks. However, most existing black-box jailbreak attacks rely on hand-cr…
cs.CR2022★ 15 cited
VeriFi: Towards Verifiable Federated Unlearning
Xiangshan Gao, Xingjun Ma, Jingyi Wang +5
Federated learning (FL) is a collaborative learning paradigm where participants jointly train a powerful model without sharing their private data. One desirable property for FL is…
cs.LG2020★ 20 cited
TrojanZoo: Towards Unified, Holistic, and Practical Evaluation of Neural Backdoors
Ren Pang, Zheng Zhang, Xiangshan Gao +5
Neural backdoors represent one primary threat to the security of deep learning systems. The intensive research has produced a plethora of backdoor attacks/defenses, resulting in a…