1 citations · 2 across the 14 of their papers we have counts for
4 papers · 1 filter
Private Direct Preference Optimization for LLM Alignment
Yangfan Jiang, Fei Wei, Ergute Bao +3
Direct preference optimization (DPO) is now a standard method for aligning large language models (LLMs) using human preference data. Each DPO example contains a prompt and a pair o…
Automated Profile Inference with Language Model Agents
Yuntao Du, Zitao Li, Bolin Ding +4
Impressive progress has been made in automated problem-solving by the collaboration of large language model (LLM) based agents. However, these automated capabilities also open aven…
Understanding Byzantine Robustness in Federated Learning with A Black-box Server
Fangyuan Zhao, Yuexiang Xie, Xuebin Ren +3
Federated learning (FL) becomes vulnerable to Byzantine attacks where some of participators tend to damage the utility or discourage the convergence of the learned model via sendin…
AAA: an Adaptive Mechanism for Locally Differential Private Mean Estimation
Fei Wei, Ergute Bao, Xiaokui Xiao +2
Local differential privacy (LDP) is a strong privacy standard that has been adopted by popular software systems. The main idea is that each individual perturbs their own data local…