2 papers
cs.LG2020
Convergence Proof for Actor-Critic Methods Applied to PPO and RUDDER
Markus Holzleitner, Lukas Gruber, José Arjona-Medina +2
We prove under commonly used assumptions the convergence of actor-critic reinforcement learning algorithms, which simultaneously learn a policy function, the actor, and a value fun…
cs.LG2020
Modern Hopfield Networks and Attention for Immune Repertoire Classification
Michael Widrich, Bernhard Schäfl, Hubert Ramsauer +8
A central mechanism in machine learning is to identify, store, and recognize patterns. How to learn, access, and retrieve such patterns is crucial in Hopfield networks and the more…