1 paper
Fengyuan Cao, Zhenhua Wang
We study infinite-horizon time-inconsistent Markov decision processes with a countably infinite state space and unbounded reward functions. The reward is allowed to depend explicit…