Autocurricula and the Emergence of Innovation from Social Interaction: A Manifesto for Multi-Agent Intelligence Research
arXiv:1903.00742
Abstract
Evolution has produced a multi-scale mosaic of interacting adaptive units. Innovations arise when perturbations push parts of the system away from stable equilibria into new regimes where previously well-adapted solutions no longer work. Here we explore the hypothesis that multi-agent systems sometimes display intrinsic dynamics arising from competition and cooperation that provide a naturally emergent curriculum, which we term an autocurriculum. The solution of one social task often begets new social tasks, continually generating novel challenges, and thereby promoting innovation. Under certain conditions these challenges may become increasingly complex over time, demanding that agents accumulate ever more innovations.
16 pages, 2 figures
Cited by in corpus (18)
- Machine Culture
- Creative Problem Solving in Artificially Intelligent Agents: A Survey and Framework
- BADGER: Learning to (Learn [Learning Algorithms] through Multi-Agent Communication)
- Open Problems in Cooperative AI
- Exploration with Unreliable Intrinsic Reward in Multi-Agent Reinforcement Learning
- Variational Automatic Curriculum Learning for Sparse-Reward Cooperative Multi-Agent Problems
- Diverse Auto-Curriculum is Critical for Successful Real-World Multiagent Learning Systems
- UneVEn: Universal Value Exploration for Multi-Agent Reinforcement Learning
- Emergent Road Rules In Multi-Agent Driving Environments
- Grounding Artificial Intelligence in the Origins of Human Behavior
- Adversarial Environment Generation for Learning to Navigate the Web
- A game-theoretic analysis of networked system control for common-pool resource management using multi-agent reinforcement learning
- Learning in Matrix Games can be Arbitrarily Complex
- Neural Auto-Curricula
- Individual and Collective Autonomous Development
- Behaviour-conditioned policies for cooperative reinforcement learning tasks
- Explore and Control with Adversarial Surprise
- Meta-control of social learning strategies