1 paper
Houlin Li, Minghui Xu, Guo Xu +8
Human-in-the-loop real-world reinforcement learning enables rapid acquisition of effective robotic manipulation policies for individual tasks, often within tens of minutes. Yet it…