8 citations · 24 across the 11 of their papers we have counts for
Showing stat.MLShow all
2 papers · 1 filter
stat.ML2023
The Effective Horizon Explains Deep RL Performance in Stochastic Environments
Cassidy Laidlaw, Banghua Zhu, Stuart Russell +1
Reinforcement learning (RL) theory has largely focused on proving minimax sample complexity bounds. These require strategic exploration algorithms that use relatively limited funct…
stat.ML2021★ 3 cited
Uncertain Decisions Facilitate Better Preference Learning
Cassidy Laidlaw, Stuart Russell
Existing observational approaches for learning human preferences, such as inverse reinforcement learning, usually make strong assumptions about the observability of the human's env…