812 citations · 827 across the 13 of their papers we have counts for
3 papers · 2 filters
Useful Policy Invariant Shaping from Arbitrary Advice
Paniz Behboudian, Yash Satsangi, Matthew E. Taylor +2
Reinforcement learning is a powerful learning paradigm in which agents can learn to maximize sparse and delayed reward signals. Although RL has had many impressive successes in com…
Sample-Efficient Model-based Actor-Critic for an Interactive Dialogue Task
Katya Kudashkina, Valliappa Chockalingam, Graham W. Taylor +1
Human-computer interactive systems that rely on machine learning are becoming paramount to the lives of millions of people who use digital assistants on a daily basis. Yet, further…
Approximate exploitability: Learning a best response in large games
Finbarr Timbers, Nolan Bard, Edward Lockhart +6
Researchers have demonstrated that neural networks are vulnerable to adversarial examples and subtle environment changes, both of which one can view as a form of distribution shift…