1 citations · 1 across the 4 of their papers we have counts for
1 paper
Ayush Aniket, Arpan Chattopadhyay
We study learning in periodic Markov Decision Process(MDP), a special type of non-stationary MDP where both the state transition probabilities and reward functions vary periodicall…