arrow
Return

Multi-class indoor semantic segmentation with deep structured model

delete2017-06-08
delete10
PRE
AI
C
Chuanxia Zheng
J
Jianhua Wang
W
Weihai Chen *
吴星明 (Xingming Wu)
DOI:10.1007/s00371-017-1411-8delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Indoor semantic segmentation plays a critical role in many applications, such as intelligent robots. However, multi-class recognition is still challenging, especially for pixel-level indoor semantic labeling. In this paper, a novel deep structured model that combines the strengths of the widely used convolutional neural networks (CNNs) and recurrent neural networks (RNNs) is proposed. We first present a multi-information fusion model that utilizes the scene category information to fine-tune the fully convolutional network. Then, to refine the coarse outputs of CNN, the RNN is applied to the final CNN layer so that we can build an end-to-end trainable system. This Graph-RNN is transformed from a conditional random field based on superpixel segmentation graphical modeling that can utilize flexible contextual information of different neighboring regions. The experimental results on the recent large SUN RGB-D dataset demonstrate that the proposed model outperforms existing state-of-the-art methods on the challenging 40 dominant classes task ( mean IU accuracy and pixel accuracy). We also evaluate our model on the public NYU depth V2 dataset and achieve remarkable performance.
Keywords:
Semantic segmentation
Scene classification
Convolutional neural network
Graph-RNN
Conditional random field
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

Visual Computer cover
Visual Computer
IF:
2.9
Papers:
4.6K
Citations:
6.5K

Organization

B
Beihang University
Scholars:
5.2W
Papers: 4.1W
Citations: 37