arrow
返回

Multi-Stream Attention-Aware Graph Convolution Network for Video Salient Object Detection

delete2021-01-01
delete43
PRE
AI
M
Mingzhu Xu
P
Ping Fu
B
Bing Liu *
J
Jun-Bao Li
DOI:10.1109/TIP.2021.3070200delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Recent advances in deep convolution neural networks (CNNs) boost the development of video salient object detection (SOD), and many remarkable deep-CNNs video SOD models have been proposed. However, many existing deep-CNNs video SOD models still suffer from coarse boundaries of the salient object, which may be attributed to the loss of high-frequency information. The traditional graph-based video SOD models can preserve object boundaries well by conducting superpixels/supervoxels segmentation in advance, but they perform weaker in highlighting the whole object than the latest deep-CNNs models, limited by heuristic graph clustering algorithms. To tackle this problem, we find a new way to address this issue under the framework of graph convolution networks (GCNs), taking advantage of graph model and deep neural network. Specifically, a superpixel-level spatiotemporal graph is first constructed among multiple frame-pairs by exploiting the motion cues implied in the frame-pairs. Then the graph data is imported into the devised multi-stream attention-aware GCN, where a novel Edge-Gated graph convolution (GC) operation is proposed to boost the saliency information aggregation on the graph data. A novel attention module is designed to encode the spatiotemporal sematic information via adaptive selection of graph nodes and fusion of the static-specific and the motion-specific graph embedding. Finally, a smoothness-aware regularization term is proposed to enhance the uniformity of salient object. Graph nodes (superpixels) inherently belonging to the same class will be ideally clustered together in the learned embedding space. Extensive experiments have been conducted on three widely used datasets. Compared with fourteen state-of-the-art video SOD models, our proposed method can well retain the salient object boundaries and possess a strong learning ability, which shows that this work is a good practice for designing GCNs for video SOD.
Keyword:
Spatiotemporal phenomena
Convolution
Data models
Visualization
Adaptation models
Object segmentation
Task analysis
Video saliency detection
spatiotemporal graph model
graph convolution network
information aggregation
node-wise attention mechanism
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

IEEE Transactions on Image Processing 封面图
IEEE Transactions on Image Processing
IF:
13.7
论文数:
1.0W
被引数:
8.4W

机构

H
harbin institute of technology
学者数:
8.0W
论文数: 6.6W
被引数: 66
引用论文

引用论文

Community gardening and social cohesion: different designs, different motivations
err2015-10-29
err0
PREAI
errE. J. Veen; B. B. Bock; W. Van den Berg; A. J. Visser; J. S. C. Wiskerke
err分享
err收藏
Video Saliency Detection via Sparsity-Based Reconstruction and Propagation
err2019-10-01
err81
PREAI
errCong, Runmin; Lei, Jianjun; Fu, Huazhu; Porikli, Fatih; Huang, Qingming; Hou, Chunping
err分享
err收藏
MATNet: Motion-Attentive Transition Network for Zero-Shot Video Object Segmentation
err2020-01-01
err152
errOAAI
errZhou, Tianfei; Li, Jianwu; Wang, Shunzhou; Tao, Ran; Shen, Jianbing
err分享
err收藏
err分享
err收藏
A Systematic Review Protocol Investigating Community Gardening Impact Measures
err2019-09-16
err0
errOAAI
errJonathan Kingsley; Aisling Bailey; Nooshin Torabi; Pauline Zardo; Suzanne Mavoa; Tonia Gray; Danielle Tracey; Philip Pettitt; Nicholas Zajac; Emily Foenander
err分享
err收藏
Bilevel Feature Learning for Video Saliency Detection
err2018-12-01
err46
PREAI
errChen, Chenglizhao; Li, Shuai; Qin, Hong; Pan, Zhenkuan; Yang, Guowei
err分享
err收藏
Salient Object Detection: A Benchmark
err2015-12-01
err687
errOAAI
errBorji, Ali; Cheng, Ming-Ming; Jiang, Huaizu; Li, Jia
err分享
err收藏
学者 查看更多内容