1 paper
Jing-Cheng Shi, Yang Yu, Qing Da +2
Applying reinforcement learning in physical-world tasks is extremely challenging. It is commonly infeasible to sample a large number of trials, as required by current reinforcement…