9 citations · 9 across the 2 of their papers we have counts for
4 papers
Discriminator Contrastive Divergence: Semi-Amortized Generative Modeling by Exploring Energy of the Discriminator
Yuxuan Song, Qiwei Ye, Minkai Xu +1
Generative Adversarial Networks (GANs) have shown great promise in modeling high dimensional data. The learning objective of GANs usually minimizes some measure discrepancy, \texti…
Suphx: Mastering Mahjong with Deep Reinforcement Learning
Junjie Li, Sotetsu Koyamada, Qiwei Ye +7
Artificial Intelligence (AI) has achieved great success in many domains, and game AI is widely regarded as its beachhead since the dawn of AI. In recent years, studies on game AI h…
Learning Efficient and Effective Exploration Policies with Counterfactual Meta Policy
Ruihan Yang, Qiwei Ye, Tie-Yan Liu
A fundamental issue in reinforcement learning algorithms is the balance between exploration of the environment and exploitation of information already obtained by the agent. Especi…
Beyond Exponentially Discounted Sum: Automatic Learning of Return Function
Yufei Wang, Qiwei Ye, Tie-Yan Liu
In reinforcement learning, Return, which is the weighted accumulated future rewards, and Value, which is the expected return, serve as the objective that guides the learning of the…