1 paper
Yonatan Ashlag, Uri Koren, Mirco Mutti +3
State entropy regularization has empirically shown better exploration and sample complexity in reinforcement learning (RL). However, its theoretical guarantees have not been studie…