2 papers
cs.LG2025
DP-AdamW: Investigating Decoupled Weight Decay and Bias Correction in Private Deep Learning
Jay Chooi, Kevin Cong, Russell Li +1
As deep learning methods increasingly utilize sensitive data on a widespread scale, differential privacy (DP) offers formal guarantees to protect against information leakage during…
cs.LG2024
Generalizing Trust: Weak-to-Strong Trustworthiness in Language Models
Martin Pawelczyk, Lillian Sun, Zhenting Qi +2
The rapid proliferation of generative AI, especially large language models, has led to their integration into a variety of applications. A key phenomenon known as weak-to-strong ge…