2 papers
cs.LG2018
Natural Option Critic
Saket Tiwari, Philip S. Thomas
The recently proposed option-critic architecture Bacon et al. provide a stochastic policy gradient approach to hierarchical reinforcement learning. Specifically, they provide a way…
cs.LG2018
Hyperbolic Embeddings for Learning Options in Hierarchical Reinforcement Learning
Saket Tiwari, M. Prannoy
Hierarchical reinforcement learning deals with the problem of breaking down large tasks into meaningful sub-tasks. Autonomous discovery of these sub-tasks has remained a challengin…