arrow
Return

Shape-based 3D human action retrieval using triplet network

delete2023-08-05
delete0
PRE
AI
王辉 (Hui Wang)
Y
Yutao Wei
B
Boxu Ding
J
Jiahao Song
王正友 (Zhengyou Wang) *
DOI:10.1007/s11042-023-16211-1delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Human action retrieval is a challenging task in computer vision and computer graphics. Existing action retrieval works are mainly for the human representation of skeleton sequences or videos. However, there are fewer researches for 3D human shapes. In this paper, we propose a novel action retrieval method for 3D human point cloud sequences. Specifically, a triplet network with the margin loss is adopted to learn embedding vectors, where their Euclidean distances are close for pairs of point cloud sequences with the same actions and are far away for pairs of sequences with different actions. Given a query point cloud sequences, the retrieval results are in ascending order via the Euclidean distances of the embedding vectors. Furthermore, we also construct a 3D human action dataset, which consists of 220 classes for evaluation of action retrieval. Extensive experiments show that the proposed method is better than the existing skeleton-based methods with a 0.05 higher precision.
Keywords:
Human action retrieval
Point cloud sequences
Triplet network

Journal

Multimedia Tools and Applications cover
Multimedia Tools and Applications
IF:
3
Papers:
1.9W
Citations:
3.2W

Organization

S
Shijiazhuang Tiedao University
Scholars:
4.1K
Papers: 2.4K
Citations: 1.7K