1 paper
Ge Li, Dong Tian, Hongyi Zhou +3
This work introduces Transformer-based Off-Policy Episodic Reinforcement Learning (TOP-ERL), a novel algorithm that enables off-policy updates in the ERL framework. In ERL, policie…