1 paper
Meshal Alharbi, Mardavij Roozbehani, Munther Dahleh
The problem of sample complexity of online reinforcement learning is often studied in the literature without taking into account any partial knowledge about the system dynamics tha…