Time-Varying Formation Controllers for Unmanned Aerial Vehicles Using Deep Reinforcement Learning
arXiv:1706.01384
Abstract
We consider the problem of designing scalable and portable controllers for unmanned aerial vehicles (UAVs) to reach time-varying formations as quickly as possible. This brief confirms that deep reinforcement learning can be used in a multi-agent fashion to drive UAVs to reach any formation while taking into account optimality and portability. We use a deep neural network to estimate how good a state is, so the agent can choose actions accordingly. The system is tested with different non-high-dimensional sensory inputs without any change in the neural network architecture, algorithm or hyperparameters, just with additional training.
Cited by in corpus (4)
- Online Multi-Objective Model-Independent Adaptive Tracking Mechanism for Dynamical Systems
- VMAV-C: A Deep Attention-based Reinforcement Learning Algorithm for Model-based Control
- Towards Closing the Sim-to-Real Gap in Collaborative Multi-Robot Deep Reinforcement Learning
- Ubiquitous Distributed Deep Reinforcement Learning at the Edge: Analyzing Byzantine Agents in Discrete Action Spaces