31 citations · 49 across the 6 of their papers we have counts for
Showing cs.AIShow all
2 papers · 1 filter
cs.AI2016★ 6 cited
Nonparametric General Reinforcement Learning
Jan Leike
Reinforcement learning (RL) problems are often phrased in terms of Markov decision processes (MDPs). In this thesis we go beyond MDPs and consider RL in environments that are non-M…
cs.AI2016★ 5 cited
A Formal Solution to the Grain of Truth Problem
Jan Leike, Jessica Taylor, Benya Fallenstein
A Bayesian agent acting in a multi-agent environment learns to predict the other agents' policies if its prior assigns positive probability to them (in other words, its prior conta…