1 citations · 1 across the 1 of their papers we have counts for
1 paper · 1 filter
Cassidy Laidlaw, Banghua Zhu, Stuart Russell +1
Reinforcement learning (RL) theory has largely focused on proving minimax sample complexity bounds. These require strategic exploration algorithms that use relatively limited funct…