1 paper
Mizhaan Prajit Maniyar, Akash Mondal, Prashanth L. A. +1
We consider the problem of control in the setting of reinforcement learning (RL), where model information is not available. Policy gradient algorithms are a popular solution approa…