arrow
返回

Deformable convolutions in multi-view stereo

delete2022-02-01
delete5
PRE
AI
J
Juliano Emir Nunes Masson *
M
Marcelo R. Petry
D
Daniel Coutinho
L
Leonardo de Mello Honório
DOI:10.1016/j.imavis.2021.104369delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
The Multi-View Stereo (MVS) is a key process in the photogrammetry workflow. It is responsible for taking the camera's views and finding the maximum number of matches between the images yielding a dense point cloud of the observed scene. Since this process is based on the matching between images it greatly depends on the abil-ity of features matching throughout different images. To improve the matching performance several researchers have proposed the use of Convolutional Neural Networks (CNNs) to solve the MVS problem. Despite the progress in the MVS problem with the usage of CNNs, the Video RAM (VRAM) consumption within these approaches is usually far greater than classical methods, that rely more on RAM, which is cheaper to expand than VRAM. This work then follows the progress made in CasMVSNet in the reduction of GPU memory usage, and further study the changes in the feature extraction process. The Average Group-wise Correlation is used in the cost vol-ume generation, to reduce the number of channels in the cost volume, yielding a reduction in GPU memory usage without noticeable penalties in the result. The deformable convolutions are applied in the feature extraction net -work to augment the spatial sampling locations with learning offsets, without additional supervision, to further improve the network's ability to model transformations. The impact of these changes is measured using quanti-tative and qualitative tests using the DTU and the Tanks and Temples datasets. The modifications reduced the GPU memory usage by 32% and improved the completeness by 9% with a penalty of 6.6% in accuracy on the DTU dataset.(c) 2021 Published by Elsevier B.V.
Keyword:
Multi-view stereo
Depth map
Deep learning

期刊

Image and Vision Computing 封面图
Image and Vision Computing
IF:
4.2
论文数:
4.1K
被引数:
6.7K

机构

U
universidade federal de juiz de fora
学者数:
5.0K
论文数: 3.4K
被引数: 2
U
universidade federal de santa catarina (ufsc)
学者数:
1.5W
论文数: 1.1W
被引数: 9
I
INESC TEC
学者数:
1.5K
论文数: 1.4K
被引数: 1.7K
学者 查看更多机构
引用论文

引用论文

USE OF BIM WITH PHOTOGRAMMETRY SUPPORT IN SMALL CONSTRUCTION PROJECTS. CASE STUDY FOR COMMERCIAL FRANCHISES
err2020-06-08
err14
errOAAI
errLoredo Conde, Antonio J.; Garcia-Sanz-Calcedo, Justo; Reyes Rodriguez, Antonio M.
err分享
err收藏
err分享
err收藏
3D Correspondence and Point Projection Method for Structures Deformation Analysis
err2020-01-01
err23
errOAAI
errMelo, Aurelio G.; Pinto, Milena F.; Honorio, Leonardo M.; Dias, Felipe M.; Masson, Juliano E. N.
err分享
err收藏
Deformable 3D Convolution for Video Super-Resolution
err2020-01-01
err113
errOAAI
errYing, Xinyi; Wang, Longguang; Wang, Yingqian; Sheng, Weidong; An, Wei; Guo, Yulan
err分享
err收藏
Novel approach to enhance coastal habitat and biotope mapping with drone aerial imagery analysis
err2021-01-12
err35
errOAAI
errMonteiro, Joao Gama; Jimenez, Jesus L.; Gizzi, Francesca; Prikryl, Petr; Lefcheck, Jonathan S.; Santos, Ricardo S.; Canning-Clode, Joao
err分享
err收藏
err分享
err收藏
Twenty-Five Year Outcomes of the Lateral Tunnel Fontan Procedure
err2017-01-01
err0
PREAI
errThomas G. Wilson; William Y. Shi; Ajay J. Iyengar; David S. Winlaw; Rachael L. Cordina; Gavin R. Wheaton; Andrew Bullock; Thomas L. Gentles; Robert G. Weintraub; Robert N. Justo; Leeanne E. Grigg; Dorothy J. Radford; Yves d'Udekem
err分享
err收藏
Effect of personality traits on sensitivity, annoyance and loudness perception of low- and high-frequency noise
err2020-07-29
err0
errOAAI
errMilad Abbasi; Mohammad Osman Tokhi; Mohsen Falahati; Saeid Yazdanirad; Maryam Ghaljahi; Siavash Etemadinezhad; Roghayeh Jaffari Talaar Poshti
err分享
err收藏
MVSNet plus plus : Learning Depth-Based Attention Pyramid Features for Multi-View Stereo
err2020-01-01
err38
PREAI
errChen, Po-Heng; Yang, Hsiao-Chien; Chen, Kuan-Wen; Chen, Yong-Sheng
err分享
err收藏
学者 查看更多内容