1 paper
Jan Corazza, Daniil Kaminskyi, Simon Lutz +4
We consider reinforcement learning in environments with dynamics that undergo an irreversible phase transition governed by a hidden temporal pattern. The agent observes the base st…