1 paper
Haibo Zhao, Dian Wang, Yizhe Zhu +6
Recent advances in hierarchical policy learning highlight the advantages of decomposing systems into high-level and low-level agents, enabling efficient long-horizon reasoning and…