3 papers
stat.ML2022
Implicit Offline Reinforcement Learning via Supervised Learning
Alexandre Piche, Rafael Pardinas, David Vazquez +2
Offline Reinforcement Learning (RL) via Supervised Learning is a simple and effective way to learn robotic skills from a dataset collected by policies of different expertise levels…
cs.LG2020
Iterative Amortized Policy Optimization
Joseph Marino, Alexandre Piché, Alessandro Davide Ialongo +1
Policy networks are a central feature of deep reinforcement learning (RL) algorithms for continuous control, enabling the estimation and sampling of high-value actions. From the va…
cs.LG2018
Reward Estimation for Variance Reduction in Deep Reinforcement Learning
Joshua Romoff, Peter Henderson, Alexandre Piché +2
Reinforcement Learning (RL) agents require the specification of a reward signal for learning behaviours. However, introduction of corrupt or stochastic rewards can yield high varia…