154 citations · 157 across the 2 of their papers we have counts for
2 papers
cs.LG2019★ 154 cited
Lyapunov-based Safe Policy Optimization for Continuous Control
Yinlam Chow, Ofir Nachum, Aleksandra Faust +2
We study continuous action reinforcement learning problems in which it is crucial that the agent interacts with the environment only through safe policies, i.e.,~policies that do n…
cs.LG2018★ 3 cited
Learning to Navigate the Web
Izzeddin Gur, Ulrich Rueckert, Aleksandra Faust +1
Learning in environments with large state and action spaces, and sparse rewards, can hinder a Reinforcement Learning (RL) agent's learning through trial-and-error. For instance, fo…