arrow
Return

Contextual Action Cues from Camera Sensor for Multi-Stream Action Recognition

delete2019-03-20
delete14
delete
OA
AI
J
Jongkwang Hong
B
Bora Cho
Y
Yong Won Hong
H
Hyeran Byun *
DOI:10.3390/s19061382delete
deleteOriginal
deleteShare
deleteSave
View PDF
Abstract

Abstract

En 中文
In action recognition research, two primary types of information are appearance and motion information that is learned from RGB images through visual sensors. However, depending on the action characteristics, contextual information, such as the existence of specific objects or globally-shared information in the image, becomes vital information to define the action. For example, the existence of the ball is vital information distinguishing kicking from running. Furthermore, some actions share typical global abstract poses, which can be used as a key to classify actions. Based on these observations, we propose the multi-stream network model, which incorporates spatial, temporal, and contextual cues in the image for action recognition. We experimented on the proposed method using C3D or inflated 3D ConvNet (I3D) as a backbone network, regarding two different action recognition datasets. As a result, we observed overall improvement in accuracy, demonstrating the effectiveness of our proposed method.
Keywords:
action recognition
contextual information
multi-stream fusion
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

Sensors cover
Sensors
IF:
3.5
Papers:
7.1W
Citations:
20.9W

Organization

Y
Yonsei University
Scholars:
4.8W
Papers: 4.6W
Citations: 5.2W