45 citations · 111 across the 19 of their papers we have counts for
5 papers · 1 filter
Code World Models for General Game Playing
Wolfgang Lehrach, Daniel Hennes, Miguel Lazaro-Gredilla +13
Large Language Models (LLMs) reasoning abilities are increasingly being applied to classical board and card games, but the dominant approach -- involving prompting for direct move…
Beyond Bayes-optimality: meta-learning what you know you don't know
Jordi Grau-Moya, Grégoire Delétang, Markus Kunesch +11
Meta-training agents with memory has been shown to culminate in Bayes-optimal agents, which casts Bayes-optimality as the implicit solution to a numerical optimization problem rath…
Causal Analysis of Agent Behavior for AI Safety
Grégoire Déletang, Jordi Grau-Moya, Miljan Martic +6
As machine learning systems become more powerful they also become increasingly unpredictable and opaque. Yet, finding human-understandable explanations of how they work is essentia…
Balancing Two-Player Stochastic Games with Soft Q-Learning
Jordi Grau-Moya, Felix Leibfried, Haitham Bou-Ammar
Within the context of video games the notion of perfectly rational agents can be undesirable as it leads to uninteresting situations, where humans face tough adversarial decision m…
An Information-Theoretic Optimality Principle for Deep Reinforcement Learning
Felix Leibfried, Jordi Grau-Moya, Haitham Bou-Ammar
We methodologically address the problem of Q-value overestimation in deep reinforcement learning to handle high-dimensional state spaces efficiently. By adapting concepts from info…