29 citations · 30 across the 3 of their papers we have counts for
1 paper · 1 filter
Jason Piquenot, Maxime Bérar, Pierre Héroux +3
This paper presents Grammar Reinforcement Learning (GRL), a reinforcement learning algorithm that uses Monte Carlo Tree Search (MCTS) and a transformer architecture that models a P…