1 paper
Mudit Verma, Ayush Kharkwal, Subbarao Kambhampati
Human-in-the-loop (HiL) reinforcement learning is gaining traction in domains with large action and state spaces, and sparse rewards by allowing the agent to take advice from HiL.…