1 paper
Saksham Gautam, Lakshmi Mandal, Shalabh Bhatnagar
In this work, we consider the problem of a two-player zero-sum game. In the literature, the successive over-relaxation Q-learning algorithm has been developed and implemented, and…