arrow
返回

Video Region Annotation with Sparse Bounding Boxes

delete2022-12-14
delete0
PRE
AI
Y
Yuzheng Xu *
Y
Yang Wu
N
Nur Sabrina binti Zuraimi
S
Shohei Nobuhara
K
Ko Nishino
DOI:10.1007/s11263-022-01719-0delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Video analysis has been moving towards more detailed interpretation (e.g., segmentation) with encouraging progress. These tasks, however, increasingly rely on densely annotated training data both in space and time. Since such annotation is labor-intensive, few densely annotated video data with detailed region boundaries exist. This work aims to resolve this dilemma by learning to automatically generate region boundaries for all frames of a video from sparsely annotated bounding boxes of target regions. We achieve this with a Volumetric Graph Convolutional Network (VGCN), which learns to iteratively find keypoints on the region boundaries using the spatio-temporal volume of surrounding appearance and motion. We show that the global optimization of VGCN leads to more accurate annotation that generalizes better. Experimental results using three latest datasets (two real and one synthetic), including ablation studies, demonstrate the effectiveness and superiority of our method.
Keyword:
Video annotation
Semi-automatic annotation
Graph convolutional network
Automatic boundary finding

期刊

International Journal of Computer Vision 封面图
International Journal of Computer Vision
IF:
9.3
论文数:
3.9K
被引数:
2.8W

机构

K
Kyoto University
学者数:
5.1W
论文数: 4.6W
被引数: 6.1W
引用论文

引用论文

Judgment Capacity, Fear of Falling, and the Risk of Falls in Community-Dwelling Older Adults: The Progetto Veneto Anziani Longitudinal Study社区居住的老年人的判断能力,对跌倒的恐惧和跌倒的风险: Progetto Veneto Anziani纵向研究
err2020-06-01
err0
errOAAI
errCaterina Trevisan; Bruno M. Zanforlini; Stefania Maggi; Marianna Noale; Federica Limongi; Marina De Rui; Maria Chiara Corti; Egle Perissinotto; Anna-Karin Welmer; Enzo Manzato; Giuseppe Sergi
err分享
err收藏
InKeV
err2018-07-27
err0
PREAI
errZaafar Ahmed; Muhammad Hamad Alizai; Affan A. Syed
err分享
err收藏
err分享
err收藏
err分享
err收藏
Semantic Understanding of Scenes Through the ADE20K Dataset通过ADE20K数据集对场景进行语义理解
err2018-12-07
err934
PREAI
errZhou, Bolei; Zhao, Hang; Puig, Xavier; Xiao, Tete; Fidler, Sanja; Barriuso, Adela; Torralba, Antonio
err分享
err收藏
GAT: a Graphical Annotation Tool for semantic regions
err2009-10-27
err13
PREAI
errGiro-i-Nieto, Xavier; Camps, Neus; Marques, Ferran
err分享
err收藏
RBS/Channeling and TEM Study of Damage Buildup in Ion Bombarded GaN
err2011-07-01
err0
errOAAI
errK. Pągowska; R. Ratajczak; A. Stonert; A. Turos; L. Nowicki; N. Sathish; P. Jóźwik; A. Muecklich
err分享
err收藏
学者 查看更多内容