7 citations · 8 across the 2 of their papers we have counts for
2 papers
cs.LG2021★ 1 cited
Policy Gradient Bayesian Robust Optimization for Imitation Learning
Zaynah Javed, Daniel S. Brown, Satvik Sharma +5
The difficulty in specifying rewards for many real-world problems has led to an increased focus on learning rewards from human feedback, such as demonstrations. However, there are…
cs.LG2021★ 7 cited
Corruption-Robust Offline Reinforcement Learning
Xuezhou Zhang, Yiding Chen, Jerry Zhu +1
We study the adversarial robustness in offline reinforcement learning. Given a batch dataset consisting of tuples , an adversary is allowed to arbitrarily modify …