1 citations · 1 across the 3 of their papers we have counts for
3 papers
Information Theoretic Adversarial Training of Large Language Models
Yiwei Zhang, Jeremiah Birrell, Reza Ebrahimi +3
Large language models (LLMs) remain vulnerable to adversarial prompting despite advances in alignment and safety, often exhibiting harmful behaviors under novel attack strategies.…
An Interactive Framework for Implementing Privacy-Preserving Federated Learning: Experiments on Large Language Models
Kasra Ahmadi, Rouzbeh Behnia, Reza Ebrahimi +4
Federated learning (FL) enhances privacy by keeping user data on local devices. However, emerging attacks have demonstrated that the updates shared by users during training can rev…
Differentially Private Stochastic Gradient Descent with Fixed-Size Minibatches: Tighter RDP Guarantees with or without Replacement
Jeremiah Birrell, Reza Ebrahimi, Rouzbeh Behnia +1
Differentially private stochastic gradient descent (DP-SGD) has been instrumental in privately training deep learning models by providing a framework to control and track the priva…