Autonomous Driving with Deep Learning: A Survey of State-of-Art Technologies
arXiv:2006.06091
Abstract
Since DARPA Grand Challenges (rural) in 2004/05 and Urban Challenges in 2007, autonomous driving has been the most active field of AI applications. Almost at the same time, deep learning has made breakthrough by several pioneers, three of them (also called fathers of deep learning), Hinton, Bengio and LeCun, won ACM Turin Award in 2019. This is a survey of autonomous driving technologies with deep learning methods. We investigate the major fields of self-driving systems, such as perception, mapping and localization, prediction, planning and control, simulation, V2X and safety etc. Due to the limited space, we focus the analysis on several key areas, i.e. 2D and 3D object detection in perception, depth estimation from cameras, multiple sensor fusion on the data, feature and task level respectively, behavior modelling and prediction of vehicle driving and pedestrian trajectories.
References in corpus (24)
- YOLOv4: Optimal Speed and Accuracy of Object Detection
- Depth Map Prediction from a Single Image using a Multi-Scale Deep Network
- DeepVO: Towards End-to-End Visual Odometry with Deep Recurrent Convolutional Neural Networks
- Monocular Human Pose Estimation: A Survey of Deep Learning-based Methods
- Deep Continuous Fusion for Multi-Sensor 3D Object Detection
- Deep Learning: A Critical Appraisal
- A Survey on Neural Architecture Search
- Learning a Driving Simulator
- Optimization for deep learning: theory and algorithms
- IPOD: Intensive Point-based Object Detector for Point Cloud
- Selective Kernel Networks
- Monocular 3D Object Detection and Box Fitting Trained End-to-End Using Intersection-over-Union Loss
- Region Proposal by Guided Anchoring
- STD: Sparse-to-Dense 3D Object Detector for Point Cloud
- Situation-Aware Pedestrian Trajectory Prediction with Spatio-Temporal Attention Model
- Social-WaGDAT: Interaction-aware Trajectory Prediction via Wasserstein Graph Double-Attention Network
- RTM3D: Real-time Monocular 3D Detection from Object Keypoints for Autonomous Driving
- VectorNet: Encoding HD Maps and Agent Dynamics from Vectorized Representation
- RefinedMPL: Refined Monocular PseudoLiDAR for 3D Object Detection in Autonomous Driving
- Voxel-FPN: multi-scale voxel feature aggregation in 3D object detection from point clouds
- Segmentation is All You Need
- A Hierarchical Architecture for Sequential Decision-Making in Autonomous Driving using Deep Reinforcement Learning
- Self-Supervised Learning of Depth and Ego-motion with Differentiable Bundle Adjustment
- Class-specific Anchoring Proposal for 3D Object Recognition in LIDAR and RGB Images