7 citations · 7 across the 3 of their papers we have counts for
3 papers
Beyond Bayes-optimality: meta-learning what you know you don't know
Jordi Grau-Moya, Grégoire Delétang, Markus Kunesch +11
Meta-training agents with memory has been shown to culminate in Bayes-optimal agents, which casts Bayes-optimality as the implicit solution to a numerical optimization problem rath…
Model-Free Risk-Sensitive Reinforcement Learning
Grégoire Delétang, Jordi Grau-Moya, Markus Kunesch +4
We extend temporal-difference (TD) learning in order to obtain risk-sensitive, model-free reinforcement learning algorithms. This extension can be regarded as modification of the R…
Shaking the foundations: delusions in sequence models for interaction and control
Pedro A. Ortega, Markus Kunesch, Grégoire Delétang +16
The recent phenomenal success of language models has reinvigorated machine learning research, and large sequence models such as transformers are being applied to a variety of domai…