arrow
返回

Monocular-Based 3-D Human Pose Estimation With Refinement Block and Special Loss Function

delete2025-02-01
delete0
delete
OA
AI
T
Tsung‐Han Tsai *
Y
Yi-Jhen Luo
DOI:10.1109/JSEN.2024.3510728delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Recently, 3-D human pose estimation (HPE) from a monocular RGB image has attracted much attention following the success of a deep convolution neural network (CNN). Many algorithms take 2.5-D heatmaps as the 3-D coordinate, whose X- and Y-axes correspond to the image coordinate, and the Z-axis corresponds to the camera coordinate. Therefore, the camera matrix or the distance between the root skeleton and the camera (the ground-truth information) is usually adopted to transform the 2.5-D coordinate to 3-D space. Since 2.5-D heatmaps ignore the conversion between 2-D and 3-D positions, they lose some conversion features and limit their applicability in the real world. In this article, we present an end-to-end framework that can utilize the contextual information in RGB images to directly predict 3-D space skeletons from a monocular image. Specifically, we use the multiloss method that depends on 2-D heatmaps, volumetric heatmaps, and a refinement block to locate the root-relative 3-D human pose. Our approach takes 2-D heatmaps and volumetric heatmaps as features to compute the loss and combine the loss from relative 3-D locations to generate the total loss. The model can learn the 2-D heatmap feature and 3-D location jointly and focus on the root-relative 3-D position in the camera coordinate. The experimental result shows that our model can predict relative 3-D human pose well on Human3.6M.
Keyword:
3-D human pose estimation (HPE)
deep convolution neural network (CNN)
root-relative 3-D human pose
volumetric heatmap

期刊

IEEE Sensors Journal 封面图
IEEE Sensors Journal
IF:
4.5
论文数:
2.2W
被引数:
7.3W

机构

N
National Central University
学者数:
1.0W
论文数: 8.6K
被引数: 6.4K
引用论文

引用论文

Kinetic Mechanism of Human Histone Acetyltransferase P/CAF
err2000-09-07
err0
PREAI
errKirk G. Tanner; Michael R. Langer; John M. Denu
err分享
err收藏
Interacting Multiple Model-Based Human Pose Estimation Using a Distributed 3D Camera Network
err2019-11-15
err18
PREAI
errHe, Haoyang; Liu, Guoliang; Zhu, Xianglai; He, Li; Tian, Guohui
err分享
err收藏
Aggression攻击行为
err2007-01-01
err0
PREAI
errE.F. Coccaro; E.C. Manning
err分享
err收藏
Lessons from 30 years’ data of Korean end-stage renal disease registry, 1985–2015
err2015-09-01
err0
errOAAI
errDong-Chan Jin; Sung Ro Yun; Seoung Woo Lee; Sang Woong Han; Won Kim; Jongha Park; Yong Kyun Kim
err分享
err收藏
edge2vec: Representation learning using edge semantics for biomedical knowledge discovery
err2019-06-10
err0
errOAAI
errZheng Gao; Gang Fu; Chunping Ouyang; Satoshi Tsutsui; Xiaozhong Liu; Jeremy Yang; Christopher Gessner; Brian Foote; David Wild; Ying Ding; Qi Yu
err分享
err收藏
学者 查看更多内容