3 papers
cs.LG2026
LoRA and Privacy: When Random Projections Help (and When They Don't)
Yaxi Hu, Johanna Düngler, Bernhard Schölkopf +1
We introduce the (Wishart) projection mechanism, a randomized map of the form with and study its differential privacy properties. For ve…
cs.LG2025
Online Learning and Unlearning
Yaxi Hu, Bernhard Schölkopf, Amartya Sanyal
We formalize the problem of online learning-unlearning, where a model is updated sequentially in an online setting while accommodating unlearning requests between updates. After a…
cs.CL2025
Differentially Private Steering for Large Language Model Alignment
Anmol Goel, Yaxi Hu, Iryna Gurevych +1
Aligning Large Language Models (LLMs) with human values and away from undesirable behaviors (such as hallucination) has become increasingly important. Recently, steering LLMs towar…