arrow
返回

Deformable graph convolutional transformer for skeleton-based action recognition

delete2022-11-17
delete2
PRE
AI
S
Shuo Chen
许
许可 (Ke Xu)
B
Bo Zhu
X
Xinghao Jiang *
T
Tanfeng Sun
DOI:10.1007/s10489-022-04302-9delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
The critical problem in skeleton-based action recognition is to extract high-level semantics from dynamic changes between skeleton joints. Therefore, Graph Convolutional Networks (GCNs) are widely applied to capture the spatial-temporal information of dynamic joint coordinates by graph-based convolution. However, previous GCNS with fixed graph convolution kernel are limited to the static topology of graphs and the geometric variations of actions. Moreover, the local information of adjacent nodes of the graph is aggregated layer by layer, which increases the model complexity. In this work, a Deformable Graph Convolutional Transformer (DGT) for skeleton-based action recognition is proposed to extract adaptive features via a flexible receptive field that is learnable. In our DGT model, a multiple-input-branches (MIB) architecture is adopted to obtain multiple information, such as joints, bones, and motions. The multiple features are fused in the Transformer Classifier. Then, the Spatial-Temporal Graph Convolution units (STGC) are used to learn a preliminary feature representation indicating both spatial and temporal dependencies on the graph. Next, a Deformable spatial-temporal compound attention backbone is followed, which learns to represent a robust feature via adaptive deformable skeleton features. The adaptive representation is obtained by dynamically adjusting its receptive field owing to the offset-based convolution method. In addition, a self-attention-based transformer classifier (TC) is designed to encode the sequence of features flattened on the spatial and temporal dimensions. The fully-connected attention mechanism further helps the high-level semantic representation by focusing on essential nodes in the graph. We evaluated DGT on two challenging large-scale datasets, NTU-RGBD 60 and NTU-RGBD 120. Experiment results support the efficacy of DGT to optimize the attention for different joints adaptively. A comparable performance but much more efficient than the state-of-the-art demonstrates the effectiveness of the proposed method.
Keyword:
Action recognition
Skeleton
Graph convolution networks
Deformable
Transformer
Attention

期刊

Applied Intelligence 封面图
Applied Intelligence
IF:
3.5
论文数:
7.6K
被引数:
1.7W

机构

S
shanghai jiao tong university
学者数:
15.7W
论文数: 11.7W
被引数: 159
引用论文

引用论文

Pretransplant Culture Selects for High-quality Porcine Islets
err2006-05-01
err0
PREAI
errJosephine K.R.A. Rijkelijkhuizen; Michael P.M. van der Burg; Annemiek Töns; Onno T. Terpstra; Eelco Bouwman
err分享
err收藏
Evaluation of the Effect of Systolic Blood Pressure and Pulse Pressure on Cognitive Function: The Women's Health and Aging Study II
err2011-12-09
err0
errOAAI
errSevil Yasar; Jean Y. Ko; Stephanie Nothelle; Michelle M. Mielke; Michelle C. Carlson
err分享
err收藏
err2002-01-01
err0
PREAI
errD. I. Ismailov; E. Sh. Alekperov; F. I. Aliev
err分享
err收藏
Strictureplasty for Crohn’s disease of the small bowel in the biologic era: long-term outcomes and risk factors for recurrence
err2020-04-18
err0
PREAI
errM. Rottoli; M. Tanzanu; C. A. Manzo; M. L. Bacchi Reggiani; P. Gionchetti; F. Rizzello; L. Boschi; G. Poggioli
err分享
err收藏
Mixed graph convolution and residual transformation network for skeleton-based action recognition
err2021-05-23
err31
PREAI
errLiu, Shuhua; Bai, Xiaoying; Fang, Ming; Li, Lanting; Hung, Chih-Cheng
err分享
err收藏
学者 查看更多内容