28 citations · 37 across the 10 of their papers we have counts for
3 papers · 1 filter
Expert-augmented actor-critic for ViZDoom and Montezumas Revenge
Michał Garmulewicz, Henryk Michalewski, Piotr Miłoś
We propose an expert-augmented actor-critic algorithm, which we evaluate on two environments with sparse rewards: Montezumas Revenge and a demanding maze from the ViZDoom suite. In…
Reinforcement Learning of Theorem Proving
Cezary Kaliszyk, Josef Urban, Henryk Michalewski +1
We introduce a theorem proving algorithm that uses practically no domain heuristics for guiding its connection-style proof search. Instead, it runs many Monte-Carlo simulations gui…
Learning to Run challenge solutions: Adapting reinforcement learning methods for neuromusculoskeletal environments
Łukasz Kidziński, Sharada Prasanna Mohanty, Carmichael Ong +26
In the NIPS 2017 Learning to Run challenge, participants were tasked with building a controller for a musculoskeletal model to make it run as fast as possible through an obstacle c…