activity
20162023
most citedA Deep Reinforcement Learning Framework for the Financial Portfolio Management Problem

140 citations · 163 across the 8 of their papers we have counts for

collaborators
Showing cs.LGShow all

8 papers · 1 filter

cs.LG2023

H-GAP: Humanoid Control with a Generalist Planner

Zhengyao Jiang, Yingchen Xu, Nolan Wagener +5

Humanoid control is an important research challenge offering avenues for integration into human-centric infrastructures and enabling physics-driven humanoid animations. The dauntin…

cs.LG2023★ 1 cited

Mildly Constrained Evaluation Policy for Offline Reinforcement Learning

Linjie Xu, Zhengyao Jiang, Jinyu Wang +2

Offline reinforcement learning (RL) methodologies enforce constraints on the policy to adhere closely to the behavior policy, thereby stabilizing value learning and mitigating the…

cs.LG2023★ 2 cited

Optimal Transport for Offline Imitation Learning

Yicheng Luo, Zhengyao Jiang, Samuel Cohen +2

With the advent of large datasets, offline reinforcement learning (RL) is a promising framework for learning good decision-making policies without the need to interact with the rea…

cs.LG2022★ 3 cited

Efficient Planning in a Compact Latent Action Space

Zhengyao Jiang, Tianjun Zhang, Michael Janner +4

Planning-based reinforcement learning has shown strong performance in tasks in discrete and low-dimensional continuous action spaces. However, planning usually brings significant c…

cs.LG2022

Graph Backup: Data Efficient Backup Exploiting Markovian Transitions

Zhengyao Jiang, Tianjun Zhang, Robert Kirk +2

The successes of deep Reinforcement Learning (RL) are limited to settings where we have a large stream of online experiences, but applying RL in the data-efficient setting with lim…

cs.LG2021★ 3 cited

Grid-to-Graph: Flexible Spatial Relational Inductive Biases for Reinforcement Learning

Zhengyao Jiang, Pasquale Minervini, Minqi Jiang +1

Although reinforcement learning has been successfully applied in many domains in recent years, we still lack agents that can systematically generalize. While relational inductive b…