2 citations · 2 across the 1 of their papers we have counts for
2 papers
cs.LG2022★ 2 cited
Safe Deep RL in 3D Environments using Human Feedback
Matthew Rahtz, Vikrant Varma, Ramana Kumar +3
Agents should avoid unsafe behaviour during both training and deployment. This typically requires a simulator and a procedural specification of unsafe behaviour. Unfortunately, a s…
cs.LG2019
An Extensible Interactive Interface for Agent Design
Matthew Rahtz, James Fang, Anca D. Dragan +1
In artificial intelligence, we often specify tasks through a reward function. While this works well in some settings, many tasks are hard to specify this way. In deep reinforcement…