3 papers
stat.ML2025
Optimal Best Arm Identification under Differential Privacy
Marc Jourdan, Achraf Azize
Best Arm Identification (BAI) algorithms are deployed in data-sensitive applications, such as adaptive clinical trials or user studies. Driven by the privacy concerns of these appl…
stat.ML2025
Optimal Regret of Bernoulli Bandits under Global Differential Privacy
Achraf Azize, Yulian Wu, Junya Honda +3
As sequential learning algorithms are increasingly applied to real life, ensuring data privacy while maintaining their utilities emerges as a timely question. In this context, regr…
cs.LG2021
Conservative Optimistic Policy Optimization via Multiple Importance Sampling
Achraf Azize, Othman Gaizi
Reinforcement Learning (RL) has been able to solve hard problems such as playing Atari games or solving the game of Go, with a unified approach. Yet modern deep RL approaches are s…