124 citations · 158 across the 4 of their papers we have counts for
4 papers
Catastrophic Jailbreak of Open-source LLMs via Exploiting Generation
Yangsibo Huang, Samyak Gupta, Mengzhou Xia +2
The rapid progress in open-source large language models (LLMs) is significantly advancing AI development. Extensive efforts have been made before model release to align their behav…
Recovering Private Text in Federated Learning of Language Models
Samyak Gupta, Yangsibo Huang, Zexuan Zhong +3
Federated learning allows distributed users to collaboratively train a model while keeping each user's data private. Recently, a growing body of work has demonstrated that an eaves…
Evaluating Gradient Inversion Attacks and Defenses in Federated Learning
Yangsibo Huang, Samyak Gupta, Zhao Song +2
Gradient inversion attack (or input recovery from gradient) is an emerging threat to the security and privacy preservation of Federated learning, whereby malicious eavesdroppers or…
EMA: Auditing Data Removal from Trained Models
Yangsibo Huang, Xiaoxiao Li, Kai Li
Data auditing is a process to verify whether certain data have been removed from a trained model. A recently proposed method (Liu et al. 20) uses Kolmogorov-Smirnov (KS) distance f…