arrow
Return

Automatic Building Extraction on High-Resolution Remote Sensing Imagery Using Deep Convolutional Encoder-Decoder With Spatial Pyramid Pooling

delete2019-01-01
delete103
delete
OA
AI
Y
Yaohui Liu
L
Lutz Gross
Z
Zhiqiang Li *
X
Xiaoli Li
X
Xiwei Fan
戚文华 cover
戚文华 (Wenhua Qi)
DOI:10.1109/ACCESS.2019.2940527delete
deleteOriginal
deleteShare
deleteSave
View PDF
Abstract

Abstract

En 中文
Automatic extraction of buildings from remote sensing imagery plays a significant role in many applications, such as urban planning and monitoring changes to land cover. Various building segmentation methods have been proposed for visible remote sensing images, especially state-of-the-art methods based on convolutional neural networks (CNNs). However, high-accuracy building segmentation from high-resolution remote sensing imagery is still a challenging task due to the potentially complex texture of buildings in general and image background. Repeated pooling and striding operations used in CNNs reduce feature resolution causing a loss of detailed information. To address this issue, we propose a light-weight deep learning model integrating spatial pyramid pooling with an encoder-decoder structure. The proposed model takes advantage of a spatial pyramid pooling module to capture and aggregate multi-scale contextual information and of the ability of encoder-decoder networks to restore losses of information. The proposed model is evaluated on two publicly available datasets; the Massachusetts roads and buildings dataset and the INRIA Aerial Image Labeling Dataset. The experimental results on these datasets show qualitative and quantitative improvement against established image segmentation models, including SegNet, FCN, U-Net, Tiramisu, and FRRN. For instance, compared to the standard U-Net, the overall accuracy gain is 1.0% (0.913 vs. 0.904) and 3.6% (0.909 vs. 0.877) with a maximal increase of 3.6% in model-training time on these two datasets. These results demonstrate that the proposed model has the potential to deliver automatic building segmentation from high-resolution remote sensing images at an accuracy that makes it a useful tool for practical application scenarios.
Keywords:
Deep learning
high-resolution remote sensing imagery
building extraction
fully convolutional networks
encoder-decoder
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

IEEE Access cover
IEEE Access
IF:
3.6
Papers:
9.8W
Citations:
29.4W

Organization

C
China Earthquake Administration
Scholars:
5.1K
Papers: 3.2K
Citations: 2.5K
I
institute of geology, cea
Scholars:
506
Papers: 429
Citations: 0
U
University of Queensland
Scholars:
5.0W
Papers: 5.1W
Citations: 9.2W
researcher View more organizations