Meta-Learning through Hebbian Plasticity in Random Networks
arXiv:2007.02686
Abstract
Lifelong learning and adaptability are two defining aspects of biological agents. Modern reinforcement learning (RL) approaches have shown significant progress in solving complex tasks, however once training is concluded, the found solutions are typically static and incapable of adapting to new information or perturbations. While it is still not completely understood how biological brains learn and adapt so efficiently from experience, it is believed that synaptic plasticity plays a prominent role in this process. Inspired by this biological mechanism, we propose a search method that, instead of optimizing the weight parameters of neural networks directly, only searches for synapse-specific Hebbian learning rules that allow the network to continuously self-organize its weights during the lifetime of the agent. We demonstrate our approach on several reinforcement learning tasks with different sensory modalities and more than 450K trainable plasticity parameters. We find that starting from completely random weights, the discovered Hebbian rules enable an agent to navigate a dynamical 2D-pixel environment; likewise they allow a simulated 3D quadrupedal robot to learn how to walk while adapting to morphological damage not seen during training and in the absence of any explicit reward or error signal in less than 100 timesteps. Code is available at https://github.com/enajx/HebbianMetaLearning.
v5: Typo in initialization values corrected. v4: Typo in equation in 3.1 corrected. v3: Bug that made diagonal patterns appear has been fixed. Simulations have been re-run and plots updated. v2: Figures 1, 7 and Table 1 updated, new results on 4.1 added, typos corrected, references added
References in corpus (9)
- Neural Architecture Search with Reinforcement Learning
- Dota 2 with Large Scale Deep Reinforcement Learning
- Recurrent World Models Facilitate Policy Evolution
- Learning to reinforcement learn
- Using Fast Weights to Attend to the Recent Past
- Random feedback weights support learning in deep neural networks
- A Biologically Plausible Learning Rule for Deep Learning in the Brain
- Evolving Inborn Knowledge For Fast Adaptation in Dynamic POMDP Problems
- Augmenting Supervised Learning by Meta-learning Unsupervised Local Rules