Places: An Image Database for Deep Scene Understanding
arXiv:1610.02055
Abstract
The rise of multi-million-item dataset initiatives has enabled data-hungry machine learning algorithms to reach near-human semantic classification at tasks such as object and scene recognition. Here we describe the Places Database, a repository of 10 million scene photographs, labeled with scene semantic categories and attributes, comprising a quasi-exhaustive list of the types of environments encountered in the world. Using state of the art Convolutional Neural Networks, we provide impressive baseline performances at scene classification. With its high-coverage and high-diversity of exemplars, the Places Database offers an ecosystem to guide future progress on currently intractable visual recognition problems.
References in corpus (1)
Cited by in corpus (14)
- Wider or Deeper: Revisiting the ResNet Model for Visual Recognition
- Analysis and Optimization of Convolutional Neural Network Architectures
- Sharing Residual Units Through Collective Tensor Factorization in Deep Neural Networks
- Situation Recognition with Graph Neural Networks
- Why my photos look sideways or upside down? Detecting Canonical Orientation of Images using Convolutional Neural Networks
- WebVision Challenge: Visual Learning and Understanding With Web Data
- Active Convolution: Learning the Shape of Convolution for Image Classification
- Deep Temporal Linear Encoding Networks
- Relating Input Concepts to Convolutional Neural Network Decisions
- An Out-of-the-box Full-network Embedding for Convolutional Neural Networks
- Cross-Domain Self-supervised Multi-task Feature Learning using Synthetic Imagery
- Multi-modal Geolocation Estimation Using Deep Neural Networks
- The Compressed Model of Residual CNDS
- Generating Video Descriptions with Topic Guidance