1 paper · 1 filter
Daniel Ajeleye, Ashutosh Trivedi, Majid Zamani
Reward machines (RMs) provide a structured way to specify non-Markovian rewards in reinforcement learning (RL), thereby improving both expressiveness and programmability. Viewed mo…