1 paper
Arun Raman, Keerthan Shagrithaya, Shalabh Bhatnagar
In this paper, we use concepts from supervisory control theory of discrete event systems to propose a method to learn optimal control policies for a finite-state Markov Decision Pr…