Places205-VGGNet Models for Scene Recognition
arXiv:1508.01667
Abstract
VGGNets have turned out to be effective for object recognition in still images. However, it is unable to yield good performance by directly adapting the VGGNet models trained on the ImageNet dataset for scene recognition. This report describes our implementation of training the VGGNets on the large-scale Places205 dataset. Specifically, we train three VGGNet models, namely VGGNet-11, VGGNet-13, and VGGNet-16, by using a Multi-GPU extension of Caffe toolbox with high computational efficiency. We verify the performance of trained Places205-VGGNet models on three datasets: MIT67, SUN397, and Places205. Our trained models achieve the state-of-the-art performance on these datasets and are made public available.
2 pages
References in corpus (3)
Cited by in corpus (14)
- Hybrid CNN and Dictionary-Based Models for Scene Recognition and Domain Adaptation
- Knowledge Guided Disambiguation for Large-Scale Scene Classification with Multi-Resolution CNNs
- Locally-Supervised Deep Hybrid Model for Scene Recognition
- Power Normalizations in Fine-grained Image, Few-shot Image and Graph Classification
- Seeing with Humans: Gaze-Assisted Neural Image Captioning
- Beyond Narrative Description: Generating Poetry from Images by Multi-Adversarial Training
- Security for Machine Learning-based Systems: Attacks and Challenges during Training and Inference
- Collaborative Layer-wise Discriminative Learning in Deep Neural Networks
- Second-order Democratic Aggregation
- A General Framework for Edited Video and Raw Video Summarization
- Better Exploiting OS-CNNs for Better Event Recognition in Images
- A Deeper Look at Power Normalizations
- FAdeML: Understanding the Impact of Pre-Processing Noise Filtering on Adversarial Machine Learning
- Action Recognition with Deep Multiple Aggregation Networks