1 paper
Fang Kong, Xiangcheng Zhang, Baoxiang Wang +1
Learning Markov decision processes (MDP) in an adversarial environment has been a challenging problem. The problem becomes even more challenging with function approximation, since…