2 papers
cs.AI2019
Learning Compositional Neural Programs with Recursive Tree Search and Planning
Thomas Pierrot, Guillaume Ligner, Scott Reed +6
We propose a novel reinforcement learning algorithm, AlphaNPI, that incorporates the strengths of Neural Programmer-Interpreters (NPI) and AlphaZero. NPI contributes structural bia…
cs.LG2018
Ranked Reward: Enabling Self-Play Reinforcement Learning for Combinatorial Optimization
Alexandre Laterre, Yunguan Fu, Mohamed Khalil Jabri +6
Adversarial self-play in two-player games has delivered impressive results when used with reinforcement learning algorithms that combine deep neural networks and tree search. Algor…