1 paper
Alekh Agarwal, Tong Zhang
Provably sample-efficient Reinforcement Learning (RL) with rich observations and function approximation has witnessed tremendous recent progress, particularly when the underlying f…