From the 1 of 4 linked papers with an AI index.
4 papers
Fully Offline Reinforcement Learning
Mattie Fellows, Clarisse Wibault, Uljad Berdica +3
The paper proposes fully offline reinforcement learning methods that use Bayesian model-based techniques to learn dynamics and evaluate policies without any online interaction, ena…
DiscoGen: Procedural Generation of Algorithm Discovery Tasks in Machine Learning
Alexander D. Goldie, Zilin Wang, Adrian Hayler +17
Automating the development of machine learning algorithms has the potential to unlock new breakthroughs. However, our ability to improve and evaluate algorithm discovery systems ha…
Recurrent Structural Policy Gradient for Partially Observable Mean Field Games
Clarisse Wibault, Johannes Forkel, Sebastian Towers +9
Mean Field Games (MFGs) provide a principled framework for modelling interactions in large population systems. However, algorithmic progress has been limited since model-free metho…
Abstraction for Offline Goal-Conditioned Reinforcement Learning
Clarisse Wibault, Alexander Goldie, Antonio Villares +2
Markov Decision Processes (MDPs) often exhibit significant redundancy due to symmetries and shared structure across state-goal pairs in real-world Goal-Conditioned Reinforcement Le…