arrow
返回

Skeleton-based human action recognition using LSTM and depthwise separable convolutional neural network

delete2025-01-11
delete0
PRE
AI
H
Hoangcong Le
C
Cheng‐Kai Lu *
C
Chen‐Chien Hsu
S
Shao-Kang Huang
DOI:10.1007/s10489-024-06082-wdelete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
In the field of computer vision, the task of human action recognition (HAR) represents a challenge, due to the complexity of capturing nuanced human movements from video data. To address this issue, researchers have developed various algorithms. In this study, a novel two-stream architecture is developed that combines LSTM with a depthwise separable convolutional neural network (DSConV) and skeleton information, with the aim of enhancing the accuracy of HAR. The 3D coordinates of each joint in the skeleton are extracted using the Mediapipe library, and the 2D coordinates are obtained using MoveNet. The proposed method comprises two streams, called the temporal LSTM module and the joint-motion module, and was developed to overcome the limitations of prior two-stream RNN models, such as the vanishing gradient problem and the difficulty of effectively extracting temporal-spatial information. A performance evaluation on the benchmark datasets of JHMDB (73.31%), Florence-3D Action (97.67%), SBU Interaction (95.2%), and Penn Action (94.0%) showcases the effectiveness of the proposed model. A comparison with state-of-the-art methods demonstrates the superior performance of the approach on these datasets. This study contributes to advancing the field of HAR, with potential applications in surveillance and robotics.
Keyword:
Human action recognition
LSTM and DSConV architecture
3D and 2D skeleton
Overlapping technique

期刊

Applied Intelligence 封面图
Applied Intelligence
IF:
3.5
论文数:
7.6K
被引数:
1.7W

机构

N
National Taiwan Normal University
学者数:
4.8K
论文数: 4.7K
被引数: 4.4K
引用论文

引用论文

Assessing Bird Communities by Point Counts: Repeated Sessions and their Duration
err2000-12-01
err0
errOAAI
errAlberto Sorace; Marco Gustin; Enrico Calvario; Luigi Ianniello; Stefano Sarrocco; Claudio Carere
err分享
err收藏
Learning 3D Skeletal Representation From Transformer for Action Recognition从Transformer学习3D骨骼表示以进行动作识别
err2022-01-01
err10
errOAAI
errCha, Junuk; Saqlain, Muhammad; Kim, Donguk; Lee, Seungeun; Lee, Seongyeong; Baek, Seungryul
err分享
err收藏
Traversing the Aging Research and Health Equity Divide: Toward Intersectional Frameworks of Research Justice and Participation
err2021-07-29
err0
errOAAI
errAndrea Gilmore-Bykovskyi; Raina Croff; Crystal M Glover; Jonathan D Jackson; Jason Resendez; Adriana Perez; Megan Zuelsdorff; Gina Green-Harris; Jennifer J Manly
err分享
err收藏
CAM-CAN: Class activation map-based categorical adversarial network
err2023-07-01
err1
errOAAI
errBatchuluun, Ganbayar; Choi, Jiho; Park, Kang Ryoung
err分享
err收藏
学者 查看更多内容