1 paper · 1 filter
Matt Oberdorfer, Matt Abuzalaf
We present the first reinforcement-learning model to self-improve its reward-modulated training implemented through a continuously improving "intuition" neural network. An agent wa…