7 papers
Global classical solutions by transport noise for reaction-diffusion systems with entropy dissipation
Antonio Agresti, Michael Kniely, Bao Quoc Tang
The existence of global classical solutions for reaction-diffusion systems arising from chemical reaction networks remains a major open problem in the deterministic setting, especi…
Hyperagents
Jenny Zhang, Bingchen Zhao, Wannan Yang +5
Self-improving AI systems aim to reduce reliance on human engineering by learning to improve their own learning and problem-solving processes. Existing approaches to self-improveme…
Efficient Offline Reinforcement Learning: First Imitate, then Improve
Adam Jelley, Trevor McInroe, Sam Devlin +1
Supervised imitation-based approaches are often favored over off-policy reinforcement learning approaches for learning policies offline, since their straightforward optimization ob…
Aligning Agents like Large Language Models
Adam Jelley, Yuhan Cao, Dave Bignell +3
Training agents to act competently in complex 3D environments from high-dimensional visual information is challenging. Reinforcement learning is conventionally used to train such a…
Adapting Vision-Language Models for Evaluating World Models
Mariya Hendriksen, Tabish Rashid, David Bignell +5
World models - generative models that simulate environment dynamics conditioned on past observations and actions - are gaining prominence in planning, simulation, and embodied AI.…
Visual Encoders for Data-Efficient Imitation Learning in Modern Video Games
Lukas Schäfer, Logan Jones, Anssi Kanervisto +7
Video games have served as useful benchmarks for the decision-making community, but going beyond Atari games towards modern games has been prohibitively expensive for the vast majo…