1 paper
Ruiqi Zhu, Tianhong Dai, Oya Celiktutan
Training a robotic policy from scratch using deep reinforcement learning methods can be prohibitively expensive due to sample inefficiency. To address this challenge, transferring…