返回
Object Pose Estimation and Feature Extraction Based on PVNet
DOI:10.1109/ACCESS.2022.3223695.png)
摘要
En 中文
In traditional industrial fields, a robot arm is usually used for high-precision or highly repetitive movements, but now, with the development of three-dimensional (3D) stereo machine vision in smart manufacturing, the smart factory has moved toward the development of a robot arm combined with the image recognition technology. Currently, in the manufacturing industry, most of the images for computing are obtained using two-dimensional (2D) machine vision; here, the 2D advantage is that the camera lens can obtain the simulation of the plane color pixel, but the disadvantage is that it cannot obtain the real space depth distance information, resulting in a more accurate analysis of the workpiece position and features. Therefore, in this study, a pixel-wise voting network (PVNet)-based object pose estimation and feature extraction was developed to perform more diverse object testing for the considered network model. Unlike other workpiece picking systems for smart manufacturing, most of the systems today still framed the workpiece in two dimensions only, but the approach proposed in this paper framed the workpiece pose in three dimensions. Thus, the network could successfully predict the pose even when the workpiece was obscured or the image was not fully captured. The images were input into the neural network by means of supervised learning, and training was performed using transformation matrices between multi-angle images of the artifacts and feature points extracted from the 3D models. The results of this study revealed the pose estimation results of various objects at different viewing angles and proposed the feature gripping strategy for the robot arm to follow this process in the future.
Keyword:
3D stereo machine vision
smart manufacturing
pose estimation
feature extraction
期刊
IF:
3.6
论文数:
9.8W
被引数:
29.4W
机构
引用论文
EUV emission spectra in collisions of multiply charged Sn ions with He and Xe多电荷Sn离子与He和Xe碰撞中的EUV发射光谱
Detecting Object Surface Keypoints From a Single RGB Image via Deep Learning Network for 6-DoF Pose Estimation
IEEE ACCESS
IF3.6
Invariant object recognition is a personalized selection of invariant features in humans, not simply explained by hierarchical feed-forward vision models
SCIENTIFIC REPORTS
IF3.9
DPRNet: Deep 3D Point Based Residual Network for Semantic Segmentation and Classification of 3D Point CloudsDPRNet: 基于深度3D点的残差网络,用于3D点云的语义分割和分类
IEEE ACCESS
IF3.6
SAR Image Change Detection via Spatial Metric Learning With an Improved Mahalanobis Distance基于改进马氏距离的空间度量学习SAR图像变化检测

