1 paper · 1 filter
C. F. Maximilian Nagy, Onur Celik, Emiliyan Gospodinov +4
Action chunking improves exploration and accelerates value propagation in long-horizon reinforcement learning, but naively applying off-policy methods to the temporally extended ac…