返回
Timed-image based deep learning for action recognition in video sequences
DOI:10.1016/j.patcog.2020.107353.png)
摘要
En 中文
The paper addresses two issues relative to machine learning on 2D + X data volumes, where 2D refers to image observation and X denotes a variable that can be associated with time, depth, wavelength, etc. The first issue addressed is conditioning these structured volumes for compatibility with respect to convolutional neural networks operating on 2D image file formats. The second issue is associated with sensitive action detection in the 2D + Time case (video clips and image time series). For the data conditioning issue, the paper first highlights that referring 2D spatial convolution to its 1D Hilbert based instance is highly accurate for information compressibility upon tight frames of convolutional networks. As a consequence of this compressibility, the paper proposes converting the 2D + X data volume into a single meta-image file format, prior to machine learning frameworks. This conversion is such that any 2D frame of the 2D + X data is reshaped as a 1D array indexed by a Hilbert space-filling curve and the third variable X of the initial file format becomes the second variable in the meta-image format. For the sensitive action recognition issue, the paper provides: (i) a 3 category video database involving non-violent, moderate and extreme violence actions; (ii) the conversion of this database into a timed meta-image database from the 2D + Time to 2D conditioning stage described above and (iii) outstanding 2-level and 3-level violence classification results from deep convolutional neural networks operating on meta-image databases. (C) 2020 Elsevier Ltd. All rights reserved.
Keyword:
Data conditioning
Video analysis
Deep learning
Convolution frames
Hilbert space-filling curve
Action recognition
Violence detection
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
IF:
7.6
论文数:
1.3W
被引数:
4.5W
机构
引用论文
Learning principal orientations and residual descriptor for action recognition
PATTERN RECOGNITION
IF7.6
Learning motion representation for real-time spatio-temporal action localization实时时空动作定位的运动表示学习
PATTERN RECOGNITION
IF7.6
Action Recognition in Video Sequences using Deep Bi-Directional LSTM With CNN Features基于CNN特征的深度双向LSTM视频序列动作识别
IEEE ACCESS
IF3.6
Spatio-temporal deformable 3D ConvNets with attention for action recognition用于动作识别的具有注意的时空可变形3D ConvNets
PATTERN RECOGNITION
IF7.6
没有更多内容

