arrow
返回

CSR-Net plus plus : Rethinking Context Structure Representation Learning for Feature Matching

delete2024-01-01
delete0
PRE
AI
X
X. Chen
J
Jiaxuan Chen *
DOI:10.1109/TGRS.2024.3431008delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Seeking good feature correspondences between two remote sensing (RS) images is an essential and important problem in the fields of RS and photogrammetry. Traditional approaches often necessitate a predefined geometric transformation model or additional manually crafted descriptors, significantly constraining the versatility. In this work, we adopt the recent context structure representation network (CSR-Net), which has shown promising performance in general feature matching problems, and propose modifications, named CSR-Net++, to overcome its main limitations. Specifically, CSR-Net is combined with a PointNet-like geometry estimator, which is sensitive to large deformations, for global preregistration. In addition, CSR-Net learns local consensus representation through a fixed-size grid, leading to limited space-aware capacities due to grid pixelwise max-pooling operations. To tackle the abovementioned limitations, we first introduce a pruning layer for matching guided by global consensus, as opposed to relying on a geometric estimator. In addition, for directly learning consensus representation from points, we propose a modified context structure representation (CSR) learning module including an independent spatial location stream and a stand-alone visual stream (VS). This decomposition separates local consensus into positional consensus and visual consensus. The proposed dual-stream representation learning not only avoids the introduction of grid anchors but also provides visual contextual priors. To demonstrate the robustness and versatility of our CSR-Net++, we conducted comprehensive experiments using diverse sets of real image pairs for general feature matching. The results demonstrate the superiority of our CSR-Net++ in most matching scenarios, achieving a 0.47%-4.70% improvement in F-score for multimodal images over existing leading methods.
Keyword:
Deformation
Representation learning
Estimation
Visualization
Task analysis
Image matching
Sensors
Deep learning
image registration
mismatch removal (MR)
remote sensing (RS) image matching
representation learning

期刊

IEEE Transactions on Geoscience and Remote Sensing 封面图
IEEE Transactions on Geoscience and Remote Sensing
IF:
8.6
论文数:
2.1W
被引数:
10.7W

机构

C
china agricultural university
学者数:
5.1W
论文数: 3.0W
被引数: 43
Z
zhejiang university
学者数:
17.7W
论文数: 12.1W
被引数: 152
引用论文

引用论文

err分享
err收藏
Communal Strength Norms in the United States and Egypt
err2013-06-28
err0
errOAAI
errSherri P. Pataki; Safia Fathelbab; Margaret S. Clark; Catharine H. Malinowski
err分享
err收藏
LMR: Learning a Two-Class Classifier for Mismatch Removal
err2019-08-01
err187
PREAI
errMa, Jiayi; Jiang, Xingyu; Jiang, Junjun; Zhao, Ji; Guo, Xiaojie
err分享
err收藏
Brain network modularity predicts cognitive training-related gains in young adults
err2019-08-01
err0
errOAAI
errPauline L. Baniqued; Courtney L. Gallen; Michael B. Kranz; Arthur F. Kramer; Mark D'Esposito
err分享
err收藏
Epidemiology of Gastric Cancer in Chile: II - Nitrate Exposures and Stomach Cancer Frequency
err1981-01-01
err0
errOAAI
errROLANDO ARMIJO; ADA GONZALEZ; MARCIAL ORELLANA; ANNE H COULSON; JAMES W SAYRE; ROGER DETELS
err分享
err收藏
学者 查看更多内容