11 citations · 23 across the 5 of their papers we have counts for
3 papers · 1 filter
Deep Quality-Value (DQV) Learning
Matthia Sabatelli, Gilles Louppe, Pierre Geurts +1
We introduce a novel Deep Reinforcement Learning (DRL) algorithm called Deep Quality-Value (DQV) Learning. DQV uses temporal-difference learning to train a Value neural network and…
Sampled Policy Gradient for Learning to Play the Game Agar.io
Anton Orell Wiehe, Nil Stolt Ansó, Madalina M. Drugan +1
In this paper, a new offline actor-critic learning algorithm is introduced: Sampled Policy Gradient (SPG). SPG samples in the action space to calculate an approximated policy gradi…
Comparing Generative Adversarial Network Techniques for Image Creation and Modification
Mathijs Pieters, Marco Wiering
Generative adversarial networks (GANs) have demonstrated to be successful at generating realistic real-world images. In this paper we compare various GAN techniques, both supervise…