1 paper · 1 filter
Kaustubh Mani, Vincent Mai, Charlie Gauthier +3
Reinforcement learning algorithms typically necessitate extensive exploration of the state space to find optimal policies. However, in safety-critical applications, the risks assoc…