45 citations · 138 across the 16 of their papers we have counts for
1 paper · 1 filter
Philip S. Thomas, Emma Brunskill
In this paper we present a new way of predicting the performance of a reinforcement learning policy given historical data that may have been generated by a different policy. The ab…