3 papers
cs.AI2020
How to Learn a Useful Critic? Model-based Action-Gradient-Estimator Policy Optimization
Pierluca D'Oro, Wojciech Jaśkowski
Deterministic-policy actor-critic algorithms for continuous control improve the actor by plugging its actions into the critic and ascending the action-value gradient, which is obta…
cs.LG2019
Artificial Intelligence for Prosthetics - challenge solutions
Łukasz Kidziński, Carmichael Ong, Sharada Prasanna Mohanty +47
In the NeurIPS 2018 Artificial Intelligence for Prosthetics challenge, participants were tasked with building a controller for a musculoskeletal model with a goal of matching a giv…
cs.LG2018
Model-Based Active Exploration
Pranav Shyam, Wojciech Jaśkowski, Faustino Gomez
Efficient exploration is an unsolved problem in Reinforcement Learning which is usually addressed by reactively rewarding the agent for fortuitously encountering novel situations.…