activity
20152020
most citedLearning Belief Representations for Imitation Learning in POMDPs

13 citations · 35 across the 11 of their papers we have counts for

collaborators

18 papers

cs.LG20203 cited

Harnessing Distribution Ratio Estimators for Learning Agents with Quality and Diversity

Tanmay Gangwani, Jian Peng, Yuan Zhou

Quality-Diversity (QD) is a concept from Neuroevolution with some intriguing applications to Reinforcement Learning. It facilitates learning a population of agents where each membe…

cs.LG2020

Off-Policy Interval Estimation with Lipschitz Value Iteration

Ziyang Tang, Yihao Feng, Na Zhang +2

Off-policy evaluation provides an essential tool for evaluating the effects of different policies or treatments using only observed data. When applied to high-stakes scenarios such…

astro-ph.CO20204 cited

Hunting for Dark Matter Subhalos in Strong Gravitational Lensing with Neural Networks

Joshua Yao-Yu Lin, Hang Yu, Warren Morningstar +2

Dark matter substructures are interesting since they can reveal the properties of dark matter. Collisionless N-body simulations of cold dark matter show more substructures compared…

cs.LG2020

Learning Guidance Rewards with Trajectory-space Smoothing

Tanmay Gangwani, Yuan Zhou, Jian Peng

Long-term temporal credit assignment is an important challenge in deep reinforcement learning (RL). It refers to the ability of the agent to attribute actions to consequences that…

cs.LG20201 cited

Efficient Competitive Self-Play Policy Optimization

Yuanyi Zhong, Yuan Zhou, Jian Peng

Reinforcement learning from self-play has recently reported many successes. Self-play, where the agents compete with themselves, is often used to generate training data for iterati…

stat.ML202010 cited

State-only Imitation with Transition Dynamics Mismatch

Tanmay Gangwani, Jian Peng

Imitation Learning (IL) is a popular paradigm for training agents to achieve complicated goals by leveraging expert behavior, rather than dealing with the hardships of designing a…