arrow
返回

Fully automatic person segmentation in unconstrained video using spatio-temporal conditional random fields

delete2016-07-01
delete8
PRE
AI
C
Chetan Bhole *
C
Christopher Pal
DOI:10.1016/j.imavis.2016.04.007delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
The segmentation of objects and people in particular is an important problem in computer vision. In this paper, we focus on automatically segmenting a person from challenging video sequences in which we place no constraint on camera viewpoint, camera motion or the movements of a person in the scene. Our approach uses the most confident predictions from a pose detector as a form of anchor or keyframe stick figure prediction which helps guide the segmentation of other more challenging frames in the video. Since even state of the art pose detectors are unreliable on many frames especially given that we are interested in segmentations with no camera or motion constraints only the poses or stick figure predictions for frames with the highest confidence in a localized temporal region anchor further processing. The stick figure predictions within confident keyframes are used to extract color, position and optical flow features. Multiple conditional random fields (CRFs) are used to process blocks of video in batches, using a two dimensional CRF for detailed keyframe segmentation as well as 3D CRFs for propagating segmentations to the entire sequence of frames belonging to batches. Location information derived from the pose is also used to refine the results. Importantly, no hand labeled training data is required by our method. We discuss the use of a continuity method that reuses learnt parameters between batches of frames and show how pose predictions can also be improved by our model. We provide an extensive evaluation of our approach, comparing it with a variety of alternative grab cut based methods and a prior state of the art method. We also release our evaluation data to the community to facilitate further experiments. We find that our approach yields state of the art qualitative and quantitative performance compared to prior work and more heuristic alternative approaches. (C) 2016 Elsevier B.V. All rights reserved.
Keyword:
Person segmentation
Video segmentation
Conditional random field
Optical flow
Fully automatic
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Image and Vision Computing 封面图
Image and Vision Computing
IF:
4.2
论文数:
4.1K
被引数:
6.7K

机构

U
universite de montreal
学者数:
4.6W
论文数: 3.8W
被引数: 46
U
University of Rochester
学者数:
2.6W
论文数: 2.1W
被引数: 2.2W
引用论文

引用论文

err分享
err收藏
Simultaneous segmentation and pose estimation of humans using dynamic graph cuts
err2008-01-10
err96
PREAI
errKohli, Pushmeet; Rihan, Jonathan; Bray, Matthieu; Torr, Philip H. S.
err分享
err收藏
Video object cut and paste
err2005-07-01
err203
PREAI
errLi, Y; Sun, J; Shum, HY
err分享
err收藏