Spatial As Deep: Spatial CNN for Traffic Scene Understanding
arXiv:1712.06080
Abstract
Convolutional neural networks (CNNs) are usually built by stacking convolutional operations layer-by-layer. Although CNN has shown strong capability to extract semantics from raw pixels, its capacity to capture spatial relationships of pixels across rows and columns of an image is not fully explored. These relationships are important to learn semantic objects with strong shape priors but weak appearance coherences, such as traffic lanes, which are often occluded or not even painted on the road surface as shown in Fig. 1 (a). In this paper, we propose Spatial CNN (SCNN), which generalizes traditional deep layer-by-layer convolutions to slice-byslice convolutions within feature maps, thus enabling message passings between pixels across rows and columns in a layer. Such SCNN is particular suitable for long continuous shape structure or large objects, with strong spatial relationship but less appearance clues, such as traffic lanes, poles, and wall. We apply SCNN on a newly released very challenging traffic lane detection dataset and Cityscapse dataset. The results show that SCNN could learn the spatial relationship for structure output and significantly improves the performance. We show that SCNN outperforms the recurrent neural network (RNN) based ReNet and MRF+CNN (MRFNet) in the lane detection dataset by 8.7% and 4.6% respectively. Moreover, our SCNN won the 1st place on the TuSimple Benchmark Lane Detection Challenge, with an accuracy of 96.53%.
Accepted to AAAI 2018
Cited by in corpus (47)
- Gen-LaneNet: A Generalized and Scalable Approach for 3D Lane Detection
- Learning Lightweight Lane Detection CNNs by Self Attention Distillation
- Key Points Estimation and Point Instance Segmentation Approach for Lane Detection
- Lane Detection Model Based on Spatio-Temporal Network With Double Convolutional Gated Recurrent Units
- A Hybrid Spatial-temporal Deep Learning Architecture for Lane Detection
- LDNet: End-to-End Lane Marking Detection Approach Using a Dynamic Vision Sensor
- 3D-LaneNet+: Anchor Free Lane Detection using a Semi-Local Representation
- RESA: Recurrent Feature-Shift Aggregator for Lane Detection
- Computing Systems for Autonomous Driving: State-of-the-Art and Challenges
- Dirty Road Can Attack: Security of Deep Learning based Automated Lane Centering under Physical-World Attack
- Advances in centerline estimation for autonomous lateral control
- CurveLane-NAS: Unifying Lane-Sensitive Architecture Search and Adaptive Point Blending
- SwiftLane: Towards Fast and Efficient Lane Detection
- Label-guided Attention Distillation for Lane Segmentation
- Driving Datasets Literature Review
- Structure-Aware Network for Lane Marker Extraction with Dynamic Vision Sensor
- CondLaneNet: a Top-to-down Lane Detection Framework Based on Conditional Convolution
- A Relation-Augmented Fully Convolutional Network for Semantic Segmentation in Aerial Scenes
- Additive Noise Annealing and Approximation Properties of Quantized Neural Networks
- Is it Safe to Drive? An Overview of Factors, Challenges, and Datasets for Driveability Assessment in Autonomous Driving
- Mcity Data Collection for Automated Vehicles Study
- Ultra Fast Structure-aware Deep Lane Detection
- Heatmap-based Vanishing Point boosts Lane Detection
- End-to-end Lane Shape Prediction with Transformers
- Inter-Region Affinity Distillation for Road Marking Segmentation
- VEGA: Towards an End-to-End Configurable AutoML Pipeline
- Semi-Local 3D Lane Detection and Uncertainty Estimation
- On Robustness of Lane Detection Models to Physical-World Adversarial Attacks in Autonomous Driving
- TextRay: Contour-based Geometric Modeling for Arbitrary-shaped Scene Text Detection
- ContinuityLearner: Geometric Continuity Feature Learning for Lane Segmentation
- 3D-LaneNet: End-to-End 3D Multiple Lane Detection
- Two at Once: Enhancing Learning and Generalization Capacities via IBN-Net
- Structure Guided Lane Detection
- Evaluating Computer Vision Techniques for Urban Mobility on Large-Scale, Unconstrained Roads
- SUPER: A Novel Lane Detection System
- Hyperspectral Image Classification with Spatial Consistence Using Fully Convolutional Spatial Propagation Network
- End-to-End Monocular Vanishing Point Detection Exploiting Lane Annotations
- Using DP Towards A Shortest Path Problem-Related Application
- A novel multimodal fusion network based on a joint coding model for lane line segmentation
- Prediction of Lane Number Using Results From Lane Detection
- Semi-supervised lane detection with Deep Hough Transform
- Hybrid tracker based optimal path tracking system for complex road environments for autonomous driving
- Multi-lane Detection Using Instance Segmentation and Attentive Voting
- Harmonious Semantic Line Detection via Maximal Weight Clique Selection
- Downhole Track Detection via Multiscale Conditional Generative Adversarial Nets
- Isometric Graph Neural Networks
- Phase Space Reconstruction Network for Lane Intrusion Action Recognition