activity
20242026
collaborators

7 papers

math.AP2026

Global classical solutions by transport noise for reaction-diffusion systems with entropy dissipation

Antonio Agresti, Michael Kniely, Bao Quoc Tang

The existence of global classical solutions for reaction-diffusion systems arising from chemical reaction networks remains a major open problem in the deterministic setting, especi…

cs.AI2026

Hyperagents

Jenny Zhang, Bingchen Zhao, Wannan Yang +5

Self-improving AI systems aim to reduce reliance on human engineering by learning to improve their own learning and problem-solving processes. Existing approaches to self-improveme…

cs.LG2025

Efficient Offline Reinforcement Learning: First Imitate, then Improve

Adam Jelley, Trevor McInroe, Sam Devlin +1

Supervised imitation-based approaches are often favored over off-policy reinforcement learning approaches for learning policies offline, since their straightforward optimization ob…

cs.LG2025

Aligning Agents like Large Language Models

Adam Jelley, Yuhan Cao, Dave Bignell +3

Training agents to act competently in complex 3D environments from high-dimensional visual information is challenging. Reinforcement learning is conventionally used to train such a…

cs.LG2025

Adapting Vision-Language Models for Evaluating World Models

Mariya Hendriksen, Tabish Rashid, David Bignell +5

World models - generative models that simulate environment dynamics conditioned on past observations and actions - are gaining prominence in planning, simulation, and embodied AI.…

cs.LG2025

Visual Encoders for Data-Efficient Imitation Learning in Modern Video Games

Lukas Schäfer, Logan Jones, Anssi Kanervisto +7

Video games have served as useful benchmarks for the decision-making community, but going beyond Atari games towards modern games has been prohibitively expensive for the vast majo…