4 papers
Regret Tail Characterization of Optimal Bandit Algorithms with Generic Rewards
Subhodip Panda, Shubhada Agrawal
We study the tail behavior of regret in stochastic multi-armed bandits for algorithms that are asymptotically optimal in expectation. While minimizing expected regret is the classi…
f-INE: A Hypothesis Testing Framework for Estimating Influence under Training Randomness
Subhodip Panda, Dhruv Tarsadiya, Shashwat Sourav +2
Influence estimation methods promise to explain and debug machine learning by estimating the impact of individual samples on the final model. Yet, existing methods collapse under t…
Unlearning in Diffusion models under Data Constraints: A Variational Inference Approach
Subhodip Panda, Varun M S, Shreyans Jain +2
For a responsible and safe deployment of diffusion models in various domains, regulating the generated outputs from these models is desirable because such models could generate und…
Adapt then Unlearn: Exploring Parameter Space Semantics for Unlearning in Generative Adversarial Networks
Piyush Tiwary, Atri Guha, Subhodip Panda +1
Owing to the growing concerns about privacy and regulatory compliance, it is desirable to regulate the output of generative models. To that end, the objective of this work is to pr…