1 paper · 1 filter
Tobias Schmähling, Matthias Burkhardt, Tobias Windisch
We propose a data augmentation method for offline reinforcement learning, motivated by active positioning problems. Particularly, our approach enables the training of off-policy mo…