Rethinking Convolutional Semantic Segmentation Learning
arXiv:1710.07991
Abstract
Deep convolutional semantic segmentation (DCSS) learning doesn't converge to an optimal local minimum with random parameters initializations; a pre-trained model on the same domain becomes necessary to achieve convergence.In this work, we propose a joint cooperative end-to-end learning method for DCSS. It addresses many drawbacks with existing deep semantic segmentation learning; the proposed approach simultaneously learn both segmentation and classification; taking away the essential need of the pre-trained model for learning convergence. We present an improved inception based architecture with partial attention gating (PAG) over encoder information. The PAG also adds to achieve faster convergence and better accuracy for segmentation task. We will show the effectiveness of this learning on a diabetic retinopathy classification and segmentation dataset.
References in corpus (7)
- Very Deep Convolutional Networks for Large-Scale Image Recognition
- Efficient Inference in Fully Connected CRFs with Gaussian Edge Potentials
- Learning Deconvolution Network for Semantic Segmentation
- Wider or Deeper: Revisiting the ResNet Model for Visual Recognition
- RefineNet: Multi-Path Refinement Networks for High-Resolution Semantic Segmentation
- Large Kernel Matters -- Improve Semantic Segmentation by Global Convolutional Network
- Deep Learning: Generalization Requires Deep Compositional Feature Space Design