arrow
返回

Learning Deep Conditional Neural Network for Image Segmentation

delete2019-07-01
delete15
PRE
AI
Q
Qiurui Wang
袁
袁春 (Chun Yuan) *
刘
刘艳 (Yan Liu)
DOI:10.1109/TMM.2018.2890360delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Combining Convolutional Neural Networks (CNNs) with Conditional Random Fields (CRFs) achieves great success among recent object segmentation methods. There are two advantages by such usage. First, CNNs can extract low-level features, which are very similar to the extracted features in primates' primary visual cortex (V1). Second, CRFs can set up the relationship between input features and output labels in a direct way. In this paper, we extend the first advantage by using CNNs for low-level feature extraction and a Structured Random Forest (SRF)-based border ownership detector for high-level feature extraction, which are similar to the outputs of primates secondary visual cortex (V2). Compared to the CRF model, an improved Conditional Boltzmann Machine (CBM), which has a multi-channel visible layer, is proposed to model the relationship between predicted labels, local and global contexts of objects with multi-scale and multilevel features. Besides, our proposed CBM model is extended for object parsing by using multivisible branches instead of a single visible layer of CBM, which cannot only segment the whole body but also the parts of the body under. These visible branches use each branch for the segmentation of the whole body or one of the body parts. All branches share the same hidden layers of CBM and train the branches under an iterative way. By exploiting object parsing, the whole body segmentation performance of object is improved. To refine the segmentation output, two kinds of optimization algorithms are proposed. The superpixel-based algorithm can re-label the overlapped regions of multiple kinds of objects. The other curve correction algorithm corrects the edges of segmented object parts by using smooth edges under a curve similarity criterion. Experiments demonstrate that our models yield competitive results for object segmentation on the PASCAL VOC 2012 dataset and for object parsing on the PennFudan Pedestrian Parsing dataset, Pedestrian Parsing Surveillance Scenes dataset, Horse-Cow parsing dataset, and PASCAL Quadrupeds dataset.
Keyword:
Segmentation
object parsing
convolutional neural networks
conditional Boltzmann machines
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

IEEE Transactions on Multimedia 封面图
IEEE Transactions on Multimedia
IF:
9.7
论文数:
4.5K
被引数:
2.4W

机构

H
hong kong polytechnic university
学者数:
3.0W
论文数: 4.1W
被引数: 921
T
tsinghua university
学者数:
11.9W
论文数: 10.0W
被引数: 137
引用论文

引用论文

err分享
err收藏
Stress induced migration of grain boundaries in Cu-bicrystals
err2019-12-25
err0
PREAI
errJann-Erik Brandenburg; Markus Schoof; Dmitri A. Molodov
err分享
err收藏
Superconducting NbN-Al hybrid technology for quantum devices
err2023-01-01
err0
errOAAI
errE. Mutsenik; S. Linzen; E. Il’ichev; M. Schmelz; M. Ziegler; V. Ripka; B. Steinbach; G. Oelsner; U. Hübner; R. Stolz
err分享
err收藏
Melting-Point Estimation of Ionic Liquids by a Group Contribution Method
err2011-12-25
err0
PREAI
errClaudia L. Aguirre; Luis A. Cisternas; José O. Valderrama
err分享
err收藏
学者 查看更多内容