1 paper
Zhipeng Liang, Xiaoteng Ma, Jose Blanchet +2
To mitigate the limitation that the classical reinforcement learning (RL) framework heavily relies on identical training and test environments, Distributionally Robust RL (DRRL) ha…