1 citations · 1 across the 1 of their papers we have counts for
1 paper
Antonio Terpin, Nicolas Lanzetti, Batuhan Yardim +2
Policy Optimization (PO) algorithms have been proven particularly suited to handle the high-dimensionality of real-world continuous control tasks. In this context, Trust Region Pol…