13 citations · 35 across the 11 of their papers we have counts for
18 papers
Harnessing Distribution Ratio Estimators for Learning Agents with Quality and Diversity
Tanmay Gangwani, Jian Peng, Yuan Zhou
Quality-Diversity (QD) is a concept from Neuroevolution with some intriguing applications to Reinforcement Learning. It facilitates learning a population of agents where each membe…
Off-Policy Interval Estimation with Lipschitz Value Iteration
Ziyang Tang, Yihao Feng, Na Zhang +2
Off-policy evaluation provides an essential tool for evaluating the effects of different policies or treatments using only observed data. When applied to high-stakes scenarios such…
Hunting for Dark Matter Subhalos in Strong Gravitational Lensing with Neural Networks
Joshua Yao-Yu Lin, Hang Yu, Warren Morningstar +2
Dark matter substructures are interesting since they can reveal the properties of dark matter. Collisionless N-body simulations of cold dark matter show more substructures compared…
Learning Guidance Rewards with Trajectory-space Smoothing
Tanmay Gangwani, Yuan Zhou, Jian Peng
Long-term temporal credit assignment is an important challenge in deep reinforcement learning (RL). It refers to the ability of the agent to attribute actions to consequences that…
Efficient Competitive Self-Play Policy Optimization
Yuanyi Zhong, Yuan Zhou, Jian Peng
Reinforcement learning from self-play has recently reported many successes. Self-play, where the agents compete with themselves, is often used to generate training data for iterati…
State-only Imitation with Transition Dynamics Mismatch
Tanmay Gangwani, Jian Peng
Imitation Learning (IL) is a popular paradigm for training agents to achieve complicated goals by leveraging expert behavior, rather than dealing with the hardships of designing a…