arrow
返回

First person video summarization using different graph representations

delete2021-06-01
delete17
PRE
AI
A
Abhimanyu Sahu
A
Ananda S. Chowdhury *
DOI:10.1016/j.patrec.2021.03.013delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
First-person video summarization has emerged as an important research problem for computer vision and multimedia communities. In this paper, we show how different graph representations can be devel-oped for accurately summarizing first-person (egocentric) videos in a computationally efficient manner. Each frame in a video is first represented as a weighted graph. A shot boundary detection method us -ing graph based mutual information is developed. We next construct a weighted graph for each shot. A representative frame from each shot is selected using a graph centrality measure. A new way of charac-terizing egocentric video frames using a graph based center-surround model is shown next. Here, each representative frame is modeled as a union of a center region (graph) and a surround region (graph). By exploiting spectral measures of dissimilarity between the two (center and surround) graphs, optimal cen-ter and surround regions are determined. Optimal regions for all frames within a shot are kept the same as that of the representative frame. Center-surround differences in entropy and optical flow values along with PHOG (Pyramidal HOG) features are extracted from each frame. All frames in a video are finally represented by another weighted graph, termed as a Video Similarity Graph (VSG). The frames are clus-tered by applying a Minimum Spanning Tree (MST) based approach with a new measure for inadmissible edges. Frames closest to the centroid of each cluster are captured to build the summary. Experimental evaluation on two benchmark datasets indicate the advantage of the proposed formulation. (c) 2021 Elsevier B.V. All rights reserved.
Keyword:
First-person video
Center-surround model
Spectral graph dissimilarity
Video similarity graph
Edge inadmissibility measure
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Pattern Recognition Letters 封面图
Pattern Recognition Letters
IF:
3.3
论文数:
8.0K
被引数:
1.6W

机构

J
Jadavpur University
学者数:
7.0K
论文数: 6.4K
被引数: 5.8K
引用论文

引用论文

Formation of palladium hydride nanoparticles in Pd/C catalyst as evidenced by in situ XAS data
err2010-03-23
err0
PREAI
errA. Yu. Stakheev; I. C. Mashkovskii; O. P. Tkachenko; K. V. Klementiev; W. Grünert; G. N. Baeva; L. M. Kustov
err分享
err收藏
err分享
err收藏
VSUMM: A mechanism designed to produce static video summaries and a novel evaluation method
err2011-01-01
err471
PREAI
errFontes de Avila, Sandra Eliza; Brandao Lopes, Ana Paula; da Luz, Antonio, Jr.; Araujo, Arnaldo de Albuquerque
err分享
err收藏
Unsupervised object-level video summarization with online motion auto-encoder
err2020-02-01
err54
errOAAI
errZhang, Yujia; Liang, Xiaodan; Zhang, Dingwen; Tan, Min; Xing, Eric P.
err分享
err收藏
Spatial and temporal scoring for egocentric video summarization以自我为中心的视频摘要的时空评分
err2016-10-01
err19
PREAI
errGuo, Zhao; Gao, Lianli; Zhen, Xiantong; Zou, Fuhao; Shen, Fumin; Zheng, Kai
err分享
err收藏
学者 查看更多内容