返回
Multi-Scale Monocular Depth Estimation Based on Global Understanding
DOI:10.1109/ACCESS.2024.3382572.png)
摘要
En 中文
With the advancement of Convolutional Neural Networks, numerous convolutional neural network-based methods have been proposed for depth estimation and have achieved significant achievements. However, the repetitive convolutional layers and spatial pooling layers in these networks often lead to a reduction in spatial resolution and loss of local information, such as edge contours. To address this issue, this study presents a multi-scale monocular depth estimation model. Specifically, a Global Understanding Module was introduced on top of a generic encoder to increase the receptive field and capture contextual information. Additionally, the decoding process incorporates a Difference Module and a Multi-scale Cascade Module to guide the decoding information and refine edge contour details. Finally, extensive experiments were conducted using the KITTI and NYUv2 datasets. For the KITTI dataset, the Absolute Relative Error (Abs. Rel) was 0.057, and the Root Mean Squared Error (RMSE) was 2.415. On the NYUv2 dataset, Abs.Rel was 0.104, and RMSE was 0.380. These results indicate that the model performs well in accurately estimating depth information.
Keyword:
Convolutional neural networks
Network architecture
Transformers
Spatial resolution
depth estimation
global understanding module
difference module
cascade module
期刊
IF:
3.6
论文数:
9.8W
被引数:
29.4W
机构
引用论文
DepthFormer: Exploiting Long-range Correlation and Local Information for Accurate Monocular Depth Estimation深度形成器: 利用远程相关和局部信息进行精确的单目深度估计
Synthesis of n-type semiconducting diamond film using diphosphorus pentaoxide as the doping source以五氧化二磷为掺杂源合成n型半导体金刚石膜
The power of obfuscation techniques in malicious JavaScript code: A measurement study恶意JavaScript代码中混淆技术的功能: 一项测量研究

