3 papers
cs.LG2019
Soft Q Network
Jingbin Liu, Shuai Liu, Xinyang Gu
Deep Q Network (DQN) is a very successful algorithm, yet the inherent problem of reinforcement learning, i.e. the exploit-explore balance, remains. In this work, we introduce entro…
cs.LG2019
Policy Optimization Reinforcement Learning with Entropy Regularization
Jingbin Liu, Xinyang Gu, Shuai Liu
Entropy regularization is an important idea in reinforcement learning, with great success in recent algorithms like Soft Q Network (SQN) and Soft Actor-Critic (SAC1). In this work,…
cs.AI2019
Reinforcement learning with world model
Jingbin Liu, Xinyang Gu, Shuai Liu
Nowadays, model-free reinforcement learning algorithms have achieved remarkable performance on many decision making and control tasks, but high sample complexity and low sample eff…