1 paper
C. F. Maximilian Nagy, Onur Celik, Emiliyan Gospodinov +4
Action chunking improves exploration and accelerates value propagation in long-horizon reinforcement learning, but naively applying off-policy methods to the temporally extended ac…