5 citations · 6 across the 3 of their papers we have counts for
6 papers
Video2Skill: Adapting Events in Demonstration Videos to Skills in an Environment using Cyclic MDP Homomorphisms
Sumedh A Sontakke, Sumegh Roychowdhury, Mausoom Sarkar +3
Humans excel at learning long-horizon tasks from demonstrations augmented with textual commentary, as evidenced by the burgeoning popularity of tutorial videos online. Intuitively,…
Information-theoretic Evolution of Model Agnostic Global Explanations
Sukriti Verma, Nikaash Puri, Piyush Gupta +1
Explaining the behavior of black box machine learning models through human interpretable rules is an important research area. Recent work has focused on explaining model behavior l…
MixBoost: Synthetic Oversampling with Boosted Mixup for Handling Extreme Imbalance
Anubha Kabra, Ayush Chopra, Nikaash Puri +4
Training a classification model on a dataset where the instances of one class outnumber those of the other class is a challenging problem. Such imbalanced datasets are standard in…
Inducing Cooperative behaviour in Sequential-Social dilemmas through Multi-Agent Reinforcement Learning using Status-Quo Loss
Pinkesh Badjatiya, Mausoom Sarkar, Abhishek Sinha +4
In social dilemma situations, individual rationality leads to sub-optimal group outcomes. Several human engagements can be modeled as a sequential (multi-step) social dilemmas. How…
Explain Your Move: Understanding Agent Actions Using Specific and Relevant Feature Attribution
Nikaash Puri, Sukriti Verma, Piyush Gupta +4
As deep reinforcement learning (RL) is applied to more tasks, there is a need to visualize and understand the behavior of learned agents. Saliency maps explain agent behavior by hi…
OpticalGAN : Generative Adversarial Networks for Continuous Variable Quantum Computation
Nilay Shrivastava, Nikaash Puri, Piyush Gupta +2
We present OpticalGAN, an extension of quantum generative adversarial networks for continuous-variable quantum computation. OpticalGAN consists of photonic variational circuits com…