2 citations · 2 across the 5 of their papers we have counts for
1 paper · 2 filters
Raphael Trumpp, Denis Hoornaert, Mirco Theile +1
Residual policy learning (RPL), in which a learned policy refines a static base policy using deep reinforcement learning (DRL), has shown strong performance across various robotic…