Dynamics-Aware Quality-Diversity for Efficient Learning of Skill Repertoires
arXiv:2109.08522 · doi:10.1109/ICRA46639.2022.9811559
Abstract
Quality-Diversity (QD) algorithms are powerful exploration algorithms that allow robots to discover large repertoires of diverse and high-performing skills. However, QD algorithms are sample inefficient and require millions of evaluations. In this paper, we propose Dynamics-Aware Quality-Diversity (DA-QD), a framework to improve the sample efficiency of QD algorithms through the use of dynamics models. We also show how DA-QD can then be used for continual acquisition of new skill repertoires. To do so, we incrementally train a deep dynamics model from experience obtained when performing skill discovery using QD. We can then perform QD exploration in imagination with an imagined skill repertoire. We evaluate our approach on three robotic experiments. First, our experiments show DA-QD is 20 times more sample efficient than existing QD approaches for skill discovery. Second, we demonstrate learning an entirely new skill repertoire in imagination to perform zero-shot learning. Finally, we show how DA-QD is useful and effective for solving a long horizon navigation task and for damage adaptation in the real world. Videos and source code are available at: https://sites.google.com/view/da-qd.
References in corpus (4)
Cited by in corpus (4)
- Deep Surrogate Assisted MAP-Elites for Automated Hearthstone Deckbuilding
- Adaptive Asynchronous Control Using Meta-learned Neural Ordinary Differential Equations
- Improving the Data Efficiency of Multi-Objective Quality-Diversity through Gradient Assistance and Crowding Exploration
- Quality-Diversity Optimisation on a Physical Robot Through Dynamics-Aware and Reset-Free Learning