13 citations · 35 across the 11 of their papers we have counts for
Showing stat.MLShow all
2 papers · 1 filter
stat.ML2020★ 10 cited
State-only Imitation with Transition Dynamics Mismatch
Tanmay Gangwani, Jian Peng
Imitation Learning (IL) is a popular paradigm for training agents to achieve complicated goals by leveraging expert behavior, rather than dealing with the hardships of designing a…
stat.ML2018
Learning Self-Imitating Diverse Policies
Tanmay Gangwani, Qiang Liu, Jian Peng
The success of popular algorithms for deep reinforcement learning, such as policy-gradients and Q-learning, relies heavily on the availability of an informative reward signal at ea…