Showing cs.LGShow all
3 papers · 1 filter
cs.LG2025
Enhancing DPSGD via Per-Sample Momentum and Low-Pass Filtering
Xincheng Xu, Thilina Ranbaduge, Qing Wang +2
Differentially Private Stochastic Gradient Descent (DPSGD) is widely used to train deep neural networks with formal privacy guarantees. However, the addition of differential privac…
cs.LG2025
Divergence-Augmented Policy Optimization
Qing Wang, Yingru Li, Jiechao Xiong +1
In deep reinforcement learning, policy optimization methods need to deal with issues such as function approximation and the reuse of off-policy data. Standard policy gradient metho…
cs.LG2024
Diverse Policies Recovering via Pointwise Mutual Information Weighted Imitation Learning
Hanlin Yang, Jian Yao, Weiming Liu +13
Recovering a spectrum of diverse policies from a set of expert trajectories is an important research topic in imitation learning. After determining a latent style for a trajectory,…