arrow
Return

Keypoint-based contextual representations for hand pose estimation

delete2023-09-14
delete0
PRE
AI
李玮玮 cover
李玮玮 (Weiwei Li) *
R
Rong Du
陈曙东 cover
陈曙东 (Shudong Chen)
DOI:10.1007/s11042-023-15713-2delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Most current methods for the hand pose estimation ignore the pixel-level relationship of hand keypoints, e.g. four specific keypoints in the same finger can form a semantically continuous area at pixel level. To make full use of pixel-level semantic information extracted from the origin RGB image, we propose a novel keypoint-based contextual representation(KCR) scheme for hand pose estimation, which can leverage pixel-level continuous contextual features based on the hand structure without using any additional labeling information. To extract hand structure information from the contextual features, we creatively design a novel keypoint representation and finger representation scheme by fusing the keypoints feature in a specific group. Then, the cross-attention mechanism is used to calculate the relation between the finger representations and contextual features to improve the feature integration. The augmented feature contains more hand structure information for the final hand pose estimation. Experimental results demonstrate that our method achieves competitive performance on various 2D and 3D hand pose estimation benchmarks.
Keywords:
Hand pose estimation
Gesture recognition
Relational context
Keypoint heatmap
Object Detection

Journal

Multimedia Tools and Applications cover
Multimedia Tools and Applications
IF:
3
Papers:
1.9W
Citations:
3.2W

Organization

C
chinese academy of sciences
Scholars:
56.5W
Papers: 44.9W
Citations: 704