arrow
Return

3D-TDC: A 3D temporal dilation convolution framework for video action recognition

delete2021-08-01
delete21
delete
OA
AI
明月 (Yue Ming)
F
Fan Feng *
C
Chao Li
J
Jing‐Hao Xue
DOI:10.1016/j.neucom.2021.03.120delete
deleteOriginal
deleteShare
deleteSave
View PDF
Abstract

Abstract

En 中文
Video action recognition is a vital area of computer vision. By adding temporal dimension into convolution structure, 3D convolution neural network owns the capacity to extract spatio-temporal features from videos. However, due to computing constraints, it is hard to input the whole video into the convolution network at one time, resulting in a limited temporal receptive field of the network. To address this issue, we propose a novel 3D temporal dilation convolution (3D-TDC) framework, to extract spatio-temporal features of actions from videos. First, we deploy the 3D temporal dilation convolution as the shallow temporal compression layer, enabling an effective capture of spatio-temporal information in a larger time domain with the reduced computational load. Then, an action recognition framework is constructed by integrating two networks with different temporal receptive fields to balance the long-short time difference. We conduct extensive experiments on three widely-used public datasets (UCF-101, HMDB-51, and Kinetics-400) for performance evaluation, and the experimental results demonstrate the effectiveness of our proposed framework in video action recognition with low computational load. (c) 2021 Elsevier B.V. All rights reserved.
Keywords:
3D convolution
Temporal dilation
Action recognition
Temporal compression
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

Neurocomputing cover
Neurocomputing
IF:
6.5
Papers:
2.5W
Citations:
6.5W

Organization

B
beijing university of posts & telecommunications
Scholars:
1.4W
Papers: 1.2W
Citations: 9
U
university of london
Scholars:
21.5W
Papers: 19.7W
Citations: 305