145 citations · 145 across the 2 of their papers we have counts for
3 papers
Revisiting Gaussian mixture critics in off-policy reinforcement learning: a sample-based approach
Bobak Shahriari, Abbas Abdolmaleki, Arunkumar Byravan +6
Actor-critic algorithms that make use of distributional policy evaluation have frequently been shown to outperform their non-distributional counterparts on many challenging control…
Making Efficient Use of Demonstrations to Solve Hard Exploration Problems
Tom Le Paine, Caglar Gulcehre, Bobak Shahriari +11
This paper introduces R2D3, an agent that makes efficient use of demonstrations to solve hard exploration problems in partially observable environments with highly variable initial…
Which Learning Algorithms Can Generalize Identity-Based Rules to Novel Inputs?
Paul Tupper, Bobak Shahriari
We propose a novel framework for the analysis of learning algorithms that allows us to say when such algorithms can and cannot generalize certain patterns from training data to tes…