Scinovex
article Open AccessTop 1% cited

Places: A 10 Million Image Database for Scene Recognition

Bolei ZhouÀgata LapedrizaAditya KhoslaAude OlivaAntonio Torralba

Abstract

The rise of multi-million-item dataset initiatives has enabled data-hungry machine learning algorithms to reach near-human semantic classification performance at tasks such as visual object and scene recognition. Here we describe the Places Database, a repository of 10 million scene photographs, labeled with scene semantic categories, comprising a large and diverse list of the types of environments encountered in the world. Using the state-of-the-art Convolutional Neural Networks (CNNs), we provide scene classification CNNs (Places-CNNs) as baselines, that significantly outperform the previous approaches. Visualization of the CNNs trained on Places shows that object detectors emerge as an intermediate representation of scene classification. With its high-coverage and high-diversity of exemplars, the Places Database along with the Places-CNNs offer a novel resource to guide future progress on scene recognition problems.

Advanced Image and Video Retrieval TechniquesImage Retrieval and Classification TechniquesAdvanced Neural Network ApplicationsComputer scienceArtificial intelligenceComputer visionImage processingImage segmentationPattern recognition (psychology)Image (mathematics)Database

Funding

  • National Science Foundation
  • Toyota Research Institute
  • Office of Naval Research
Citations
3,945
FWCI
103.32
field-weighted impact
References
49
Percentile
100%
vs. same field & year
Citations per year
Cited by
Deep Learning for Generic Object Detection: A Survey
International Journal of Computer Vision · 2019 · 2,702 citations
Squeeze-and-Excitation Networks
IEEE Transactions on Pattern Analysis and Machine Intelligence · 2019 · 12,333 citations
Learning to Prompt for Vision-Language Models
International Journal of Computer Vision · 2022 · 2,413 citations
Wider or Deeper: Revisiting the ResNet Model for Visual Recognition
Pattern Recognition · 2019 · 1,584 citations
Learning without Forgetting
IEEE Transactions on Pattern Analysis and Machine Intelligence · 2017 · 3,673 citations
References
Modeling the Shape of the Scene: A Holistic Representation of the Spatial Envelope
International Journal of Computer Vision · 2001 · 6,378 citations
The Pascal Visual Object Classes (VOC) Challenge
International Journal of Computer Vision · 2009 · 19,127 citations
Long Short-Term Memory
Neural Computation · 1997 · 95,078 citations
WordNet
Communications of the ACM · 1995 · 13,991 citations
Measurement of Diversity
Nature · 1949 · 13,763 citations
Gradient-based learning applied to document recognition
Proceedings of the IEEE · 1998 · 57,014 citations
ImageNet Large Scale Visual Recognition Challenge
International Journal of Computer Vision · 2015 · 39,683 citations
ImageNet classification with deep convolutional neural networks
Communications of the ACM · 2017 · 75,550 citations
Citation Network

How this paper connects to the literature. Drag to explore, click any node to open that paper.