3 citations · 3 across the 2 of their papers we have counts for
2 papers
cs.LG2024
Toward Information Theoretic Active Inverse Reinforcement Learning
Ondrej Bajgar, Sid William Gould, Rohan Narayan Langford Mitta +3
As AI systems become increasingly autonomous, aligning their decision-making to human preferences is essential. In domains like autonomous driving or robotics, it is impossible to…
cs.LG2023★ 3 cited
Safeguarded Progress in Reinforcement Learning: Safe Bayesian Exploration for Control Policy Synthesis
Rohan Mitta, Hosein Hasanbeig, Jun Wang +3
This paper addresses the problem of maintaining safety during training in Reinforcement Learning (RL), such that the safety constraint violations are bounded at any point during le…