318 citations
- Apple (United States)US3 papers
- Carnegie Mellon UniversityUS3 papers
- Amazon (Germany)DE1 paper
- Duke UniversityUS1 paper
- Google (United States)US1 paper
- Instituto Superior TécnicoPT1 paper
- Karlsruhe Institute of TechnologyDE1 paper
- Martin Luther University Halle-WittenbergDE1 paper
- Microsoft (Finland)FI1 paper
- Microsoft Research (United Kingdom)GB1 paper
- Moscow Institute of Thermal TechnologyRU1 paper
- Princeton UniversityUS1 paper
Showing 2021 · cs.LGShow all
3 papers · 2 filters
cs.LG2021★ 12 cited
Private Adaptive Gradient Methods for Convex Optimization
Hilal Asi, John Duchi, Alireza Fallah +2
We study adaptive methods for differentially private convex optimization, proposing and analyzing differentially private variants of a Stochastic Gradient Descent (SGD) algorithm w…
cs.LG2021★ 20 cited
Uncertainty Weighted Actor-Critic for Offline Reinforcement Learning
Yue Wu, Shuangfei Zhai, Nitish Srivastava +4
Offline Reinforcement Learning promises to learn effective policies from previously-collected, static datasets without the need for exploration. However, existing Q-learning and ac…
cs.LG2021★ 1 cited
MetricOpt: Learning to Optimize Black-Box Evaluation Metrics
Chen Huang, Shuangfei Zhai, Pengsheng Guo +1
We study the problem of directly optimizing arbitrary non-differentiable task evaluation metrics such as misclassification rate and recall. Our method, named MetricOpt, operates in…