5 citations · 6 across the 2 of their papers we have counts for
3 papers
Explicit Tradeoffs between Adversarial and Natural Distributional Robustness
Mazda Moayeri, Kiarash Banihashem, Soheil Feizi
Several existing works study either adversarial or natural distributional robustness of deep neural networks separately. In practice, however, models need to enjoy both types of ro…
Admissible Policy Teaching through Reward Design
Kiarash Banihashem, Adish Singla, Jiarui Gan +1
We study reward design strategies for incentivizing a reinforcement learning agent to adopt a policy from a set of admissible policies. The goal of the reward designer is to modify…
Defense Against Reward Poisoning Attacks in Reinforcement Learning
Kiarash Banihashem, Adish Singla, Goran Radanovic
We study defense strategies against reward poisoning attacks in reinforcement learning. As a threat model, we consider attacks that minimally alter rewards to make the attacker's t…