2 citations · 3 across the 2 of their papers we have counts for
1 paper · 1 filter
Yinan Zhang, Devin Balkcom, Haoxiang Li
This paper addresses the question of how a previously available control policy πs can be used as a supervisor to more quickly and safely train a new learned control policy πL…