Recent Advances in Leveraging Human Guidance for Sequential Decision-Making Tasks
arXiv:2107.05825 · doi:10.1007/s10458-021-09514-w
Abstract
A longstanding goal of artificial intelligence is to create artificial agents capable of learning to perform tasks that require sequential decision making. Importantly, while it is the artificial agent that learns and acts, it is still up to humans to specify the particular task to be performed. Classical task-specification approaches typically involve humans providing stationary reward functions or explicit demonstrations of the desired tasks. However, there has recently been a great deal of research energy invested in exploring alternative ways in which humans may guide learning agents that may, e.g., be more suitable for certain tasks or require less human effort. This survey provides a high-level overview of five recent machine learning frameworks that primarily rely on human guidance apart from pre-specified reward functions or conventional, step-by-step action demonstrations. We review the motivation, assumptions, and implementation of each framework, and we discuss possible future research directions.
Springer journal, Autonomous Agents and Multi-Agent Systems (JAAMAS)
References in corpus (10)
- FeUdal Networks for Hierarchical Reinforcement Learning
- Trial without Error: Towards Safe Reinforcement Learning via Human Intervention
- Learning Invariant Feature Spaces to Transfer Skills with Reinforcement Learning
- Extrapolating Beyond Suboptimal Demonstrations via Inverse Reinforcement Learning from Observations
- Imitation Learning from Observations by Minimizing Inverse Dynamics Disagreement
- Digging Deeper into Egocentric Gaze Prediction
- Provably Efficient Imitation Learning from Observation Alone
- Relay Policy Learning: Solving Long-Horizon Tasks via Imitation and Reinforcement Learning
- DDCO: Discovery of Deep Continuous Options for Robot Learning from Demonstrations
- Model-based Behavioral Cloning with Future Image Similarity Learning