1 paper
Y. Cheng, P. Zhao, F. Wang +2
A reinforcement learning (RL) control policy could fail in a new/perturbed environment that is different from the training environment, due to the presence of dynamic variations. F…