3D Object Detection for Autonomous Driving: A Survey
arXiv:2106.10823 · doi:10.1016/j.patcog.2022.108796
Abstract
Autonomous driving is regarded as one of the most promising remedies to shield human beings from severe crashes. To this end, 3D object detection serves as the core basis of perception stack especially for the sake of path planning, motion prediction, and collision avoidance etc. Taking a quick glance at the progress we have made, we attribute challenges to visual appearance recovery in the absence of depth information from images, representation learning from partially occluded unstructured point clouds, and semantic alignments over heterogeneous features from cross modalities. Despite existing efforts, 3D object detection for autonomous driving is still in its infancy. Recently, a large body of literature have been investigated to address this 3D vision task. Nevertheless, few investigations have looked into collecting and structuring this growing knowledge. We therefore aim to fill this gap in a comprehensive survey, encompassing all the main concerns including sensors, datasets, performance metrics and the recent state-of-the-art detection methods, together with their pros and cons. Furthermore, we provide quantitative comparisons with the state of the art. A case study on fifteen selected representative methods is presented, involved with runtime analysis, error analysis, and robustness analysis. Finally, we provide concluding remarks after an in-depth analysis of the surveyed works and identify promising directions for future work.
The manuscript is accepted by Pattern Recognition on 14 May 2022
References in corpus (15)
- Faster R-CNN: Towards Real-Time Object Detection with Region Proposal Networks
- Inductive Representation Learning on Large Graphs
- Deep Multi-modal Object Detection and Semantic Segmentation for Autonomous Driving: Datasets, Methods, and Challenges
- A Survey of Deep Learning-based Object Detection
- What makes for effective detection proposals?
- Object Detection in 20 Years: A Survey
- Deep Continuous Fusion for Multi-Sensor 3D Object Detection
- Vehicle Detection from 3D Lidar Using Fully Convolutional Network
- Frustum ConvNet: Sliding Frustums to Aggregate Local Point-Wise Features for Amodal 3D Object Detection
- PointRGCN: Graph Convolution Networks for 3D Vehicles Detection Refinement
- Voxel-FPN: multi-scale voxel feature aggregation in 3D object detection from point clouds
- HVNet: Hybrid Voxel Network for LiDAR Based 3D Object Detection
- IoU-aware Single-stage Object Detector for Accurate Localization
- What You See is What You Get: Exploiting Visibility for 3D Object Detection
- LiDAR and Camera Calibration using Motion Estimated by Sensor Fusion Odometry
Cited by in corpus (13)
- Trajectory Prediction for Autonomous Driving: Progress, Limitations, and Future Directions
- LiDAR Spoofing Meets the New-Gen: Capability Improvements, Broken Assumptions, and New Attack Strategies
- Model-agnostic explainable artificial intelligence for object detection in image data
- Towards pedestrian head tracking: A benchmark dataset and a multi-source data fusion network
- A Multimodal Hybrid Late-Cascade Fusion Network for Enhanced 3D Object Detection
- Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities
- Composite Convolution: a Flexible Operator for Deep Learning on 3D Point Clouds
- All You Need for Object Detection: From Pixels, Points, and Prompts to Next-Gen Fusion and Multimodal LLMs/VLMs in Autonomous Vehicles
- InScope: A New Real-world 3D Infrastructure-side Collaborative Perception Dataset for Open Traffic Scenarios
- LCF3D: A Robust and Real-Time Late-Cascade Fusion Framework for 3D Object Detection in Autonomous Driving
- Benchmarking the Spatial Robustness of DNNs via Natural and Adversarial Localized Corruptions
- DPO: Dual-Perturbation Optimization for Test-time Adaptation in 3D Object Detection
- Network Optimization Aspects of Autonomous Vehicles: Challenges and Future Directions