Scinovex
article Open AccessTop 1% cited

Transferring Deep Convolutional Neural Networks for the Scene Classification of High-Resolution Remote Sensing Imagery

Remote Sensing · 2015 · Vol. 7(11) · pp. 14680–14707
Fan HuGui-Song XiaJingwen HuLiangpei Zhang

Abstract

Learning efficient image representations is at the core of the scene classification task of remote sensing imagery. The existing methods for solving the scene classification task, based on either feature coding approaches with low-level hand-engineered features or unsupervised feature learning, can only generate mid-level image features with limited representative ability, which essentially prevents them from achieving better performance. Recently, the deep convolutional neural networks (CNNs), which are hierarchical architectures trained on large-scale datasets, have shown astounding performance in object recognition and detection. However, it is still not clear how to use these deep convolutional neural networks for high-resolution remote sensing (HRRS) scene classification. In this paper, we investigate how to transfer features from these successfully pre-trained CNNs for HRRS scene classification. We propose two scenarios for generating image features via extracting CNN features from different layers. In the first scenario, the activation vectors extracted from fully-connected layers are regarded as the final image features; in the second scenario, we extract dense features from the last convolutional layer at multiple scales and then encode the dense features into global image features through commonly used feature coding approaches. Extensive experiments on two public scene classification datasets demonstrate that the image features obtained by the two proposed scenarios, even with a simple linear classifier, can result in remarkable performance and improve the state-of-the-art by a significant margin. The results reveal that the features from pre-trained CNNs generalize well to HRRS datasets and are more expressive than the low- and mid-level features. Moreover, we tentatively combine features extracted from different CNN models for better performance.

Advanced Image and Video Retrieval TechniquesRemote-Sensing Image ClassificationDomain Adaptation and Few-Shot LearningComputer scienceConvolutional neural networkArtificial intelligencePattern recognition (psychology)Classifier (UML)Contextual image classificationDeep learningTransfer of learningFeature (linguistics)Feature extraction

Funding

  • National Natural Science Foundation of China
  • Wuhan Municipal Science and Technology Bureau
Citations
1,197
FWCI
56.92
field-weighted impact
References
63
Percentile
100%
vs. same field & year
Citations per year
References
Learning representations by back-propagating errors
Nature · 1986 · 30,045 citations
Gradient-based learning applied to document recognition
Proceedings of the IEEE · 1998 · 57,014 citations
ImageNet Large Scale Visual Recognition Challenge
International Journal of Computer Vision · 2015 · 39,683 citations
Distinctive Image Features from Scale-Invariant Keypoints
International Journal of Computer Vision · 2004 · 54,768 citations
Multiresolution gray-scale and rotation invariant texture classification with local binary patterns
IEEE Transactions on Pattern Analysis and Machine Intelligence · 2002 · 15,129 citations
ImageNet classification with deep convolutional neural networks
Communications of the ACM · 2017 · 75,550 citations
Representation Learning: A Review and New Perspectives
IEEE Transactions on Pattern Analysis and Machine Intelligence · 2013 · 12,724 citations
Citation Network

How this paper connects to the literature. Drag to explore, click any node to open that paper.