8 citations · 8 across the 1 of their papers we have counts for
3 papers
cs.LG2020★ 8 cited
Representations for Stable Off-Policy Reinforcement Learning
Dibya Ghosh, Marc G. Bellemare
Reinforcement learning with function approximation can be unstable and even divergent, especially when combined with off-policy learning and Bellman updates. In deep reinforcement…
cs.LG2020
On Catastrophic Interference in Atari 2600 Games
William Fedus, Dibya Ghosh, John D. Martin +3
Model-free deep reinforcement learning is sample inefficient. One hypothesis -- speculated, but not confirmed -- is that catastrophic interference within an environment inhibits le…
cs.LG2019
Learning to Reach Goals via Iterated Supervised Learning
Dibya Ghosh, Abhishek Gupta, Ashwin Reddy +4
Current reinforcement learning (RL) algorithms can be brittle and difficult to use, especially when learning goal-reaching behaviors from sparse rewards. Although supervised imitat…