1 paper · 1 filter
Hao Shi, Xi Li
Long-horizon goal-conditioned reinforcement learning delegates control to a high-level module that proposes subgoals, but existing subgoals are implicit byproducts of value functio…