140 citations · 163 across the 8 of their papers we have counts for
8 papers · 1 filter
H-GAP: Humanoid Control with a Generalist Planner
Zhengyao Jiang, Yingchen Xu, Nolan Wagener +5
Humanoid control is an important research challenge offering avenues for integration into human-centric infrastructures and enabling physics-driven humanoid animations. The dauntin…
Mildly Constrained Evaluation Policy for Offline Reinforcement Learning
Linjie Xu, Zhengyao Jiang, Jinyu Wang +2
Offline reinforcement learning (RL) methodologies enforce constraints on the policy to adhere closely to the behavior policy, thereby stabilizing value learning and mitigating the…
Optimal Transport for Offline Imitation Learning
Yicheng Luo, Zhengyao Jiang, Samuel Cohen +2
With the advent of large datasets, offline reinforcement learning (RL) is a promising framework for learning good decision-making policies without the need to interact with the rea…
Efficient Planning in a Compact Latent Action Space
Zhengyao Jiang, Tianjun Zhang, Michael Janner +4
Planning-based reinforcement learning has shown strong performance in tasks in discrete and low-dimensional continuous action spaces. However, planning usually brings significant c…
Graph Backup: Data Efficient Backup Exploiting Markovian Transitions
Zhengyao Jiang, Tianjun Zhang, Robert Kirk +2
The successes of deep Reinforcement Learning (RL) are limited to settings where we have a large stream of online experiences, but applying RL in the data-efficient setting with lim…
Grid-to-Graph: Flexible Spatial Relational Inductive Biases for Reinforcement Learning
Zhengyao Jiang, Pasquale Minervini, Minqi Jiang +1
Although reinforcement learning has been successfully applied in many domains in recent years, we still lack agents that can systematically generalize. While relational inductive b…