1 paper
Duc Kien Doan, Bang Giang Le, Viet Cuong Ta
In safe reinforcement learning, agent needs to balance between exploration actions and safety constraints. Following this paradigm, domain transfer approaches learn a prior Q-funct…