1 citations · 1 across the 1 of their papers we have counts for
1 paper
Philipp Scholl, Felix Dietrich, Clemens Otte +1
Safe Policy Improvement (SPI) is an important technique for offline reinforcement learning in safety critical applications as it improves the behavior policy with a high probabilit…