arrow
Return

GCNDepth: Self-supervised monocular depth estimation based on graph convolutional network

delete2023-01-01
delete31
delete
OA
AI
A
Armin Masoumian *
H
Hatem A. Rashwan
S
Saddam Abdulwahab
J
Julián Cristiano
A
Asif, M. Salman
P
Puig, Domenec
DOI:10.1016/j.neucom.2022.10.073delete
deleteOriginal
deleteShare
deleteSave
View PDF
Abstract

Abstract

En 中文
Depth estimation is a challenging task of 3D reconstruction to enhance the accuracy sensing of environ-ment awareness. This work brings a new solution with improvements, which increases the quantitative and qualitative understanding of depth maps compared to existing methods. Recently, convolutional neural networks (CNN) have demonstrated their extraordinary ability to estimate depth maps from monocular videos. However, traditional CNN does not support a topological structure, and they can work only on regular image regions with determined sizes and weights. On the other hand, graph convolu-tional networks (GCN) can handle the convolution of non-Euclidean data, and they can be applied to irregular image regions within a topological structure. Therefore, to preserve object geometric appear-ances and objects locations in the scene, in this work, we aim to exploit GCN for a self-supervised monoc-ular depth estimation model. Our model consists of two parallel auto-encoder networks: the first is an auto-encoder that will depend on ResNet-50 and extract the feature from the input image and on multi-scale GCN to estimate the depth map. In turn, the second network will be used to estimate the ego-motion vector (i.e., 3D pose) between two consecutive frames based on ResNet-18. The estimated 3D pose and depth map will be used to construct the target image. A combination of loss functions related to photometric, reprojection, and smoothness is used to cope with bad depth prediction and pre-serve the discontinuities of the objects. Our method and performance are improved quantitatively and qualitatively. In particular, our method provided comparable and promising results with a high predic-tion accuracy of 89% on the publicly available KITTI dataset. Our method also offers 40% reduction in the number of trainable parameters compared to the state of the art solutions.In addition, we tested our trained model with Make3D dataset to evaluate the trained model on a new dataset with low reso-lution images. The source code is publicly available at (https://github.com/ArminMasoumian/GCNDepth. git)(c) 2022 Elsevier B.V. All rights reserved.
Keywords:
Deep learning
Graph convolutional network
Monocular depth estimation
Self -supervision
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

Neurocomputing cover
Neurocomputing
IF:
6.5
Papers:
2.5W
Citations:
6.5W

Organization

U
Universitat Rovira i Virgili
Scholars:
1.0W
Papers: 8.4K
Citations: 9.0K
University of California System cover
University of California System
Scholars:
37.5W
Papers: 33.7W
Citations: 6.6K