2 papers
cs.LG2025
Trust-Region Twisted Policy Improvement
Joery A. de Vries, Jinke He, Yaniv Oren +1
Monte-Carlo tree search (MCTS) has driven many recent breakthroughs in deep reinforcement learning (RL). However, scaling MCTS to parallel compute has proven challenging in practic…
cs.LG2025
Bayesian Meta-Reinforcement Learning with Laplace Variational Recurrent Networks
Joery A. de Vries, Jinke He, Mathijs M. de Weerdt +1
Meta-reinforcement learning trains a single reinforcement learning agent on a distribution of tasks to quickly generalize to new tasks outside of the training set at test time. Fro…