arrow
Return

CSR-Net plus plus : Rethinking Context Structure Representation Learning for Feature Matching

delete2024-01-01
delete0
PRE
AI
X
X. Chen
J
Jiaxuan Chen *
DOI:10.1109/TGRS.2024.3431008delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Seeking good feature correspondences between two remote sensing (RS) images is an essential and important problem in the fields of RS and photogrammetry. Traditional approaches often necessitate a predefined geometric transformation model or additional manually crafted descriptors, significantly constraining the versatility. In this work, we adopt the recent context structure representation network (CSR-Net), which has shown promising performance in general feature matching problems, and propose modifications, named CSR-Net++, to overcome its main limitations. Specifically, CSR-Net is combined with a PointNet-like geometry estimator, which is sensitive to large deformations, for global preregistration. In addition, CSR-Net learns local consensus representation through a fixed-size grid, leading to limited space-aware capacities due to grid pixelwise max-pooling operations. To tackle the abovementioned limitations, we first introduce a pruning layer for matching guided by global consensus, as opposed to relying on a geometric estimator. In addition, for directly learning consensus representation from points, we propose a modified context structure representation (CSR) learning module including an independent spatial location stream and a stand-alone visual stream (VS). This decomposition separates local consensus into positional consensus and visual consensus. The proposed dual-stream representation learning not only avoids the introduction of grid anchors but also provides visual contextual priors. To demonstrate the robustness and versatility of our CSR-Net++, we conducted comprehensive experiments using diverse sets of real image pairs for general feature matching. The results demonstrate the superiority of our CSR-Net++ in most matching scenarios, achieving a 0.47%-4.70% improvement in F-score for multimodal images over existing leading methods.
Keywords:
Deformation
Representation learning
Estimation
Visualization
Task analysis
Image matching
Sensors
Deep learning
image registration
mismatch removal (MR)
remote sensing (RS) image matching
representation learning

Journal

IEEE Transactions on Geoscience and Remote Sensing cover
IEEE Transactions on Geoscience and Remote Sensing
IF:
8.6
Papers:
2.1W
Citations:
10.7W

Organization

C
china agricultural university
Scholars:
5.1W
Papers: 3.0W
Citations: 43
Z
zhejiang university
Scholars:
17.7W
Papers: 12.1W
Citations: 152
Cited Papers

Cited Papers

Communal Strength Norms in the United States and Egypt
err2013-06-28
err0
errOAAI
errSherri P. Pataki; Safia Fathelbab; Margaret S. Clark; Catharine H. Malinowski
errShare
errSave
LMR: Learning a Two-Class Classifier for Mismatch Removal
err2019-08-01
err187
PREAI
errMa, Jiayi; Jiang, Xingyu; Jiang, Junjun; Zhao, Ji; Guo, Xiaojie
errShare
errSave
Brain network modularity predicts cognitive training-related gains in young adults
err2019-08-01
err0
errOAAI
errPauline L. Baniqued; Courtney L. Gallen; Michael B. Kranz; Arthur F. Kramer; Mark D'Esposito
errShare
errSave
Epidemiology of Gastric Cancer in Chile: II - Nitrate Exposures and Stomach Cancer Frequency
err1981-01-01
err0
errOAAI
errROLANDO ARMIJO; ADA GONZALEZ; MARCIAL ORELLANA; ANNE H COULSON; JAMES W SAYRE; ROGER DETELS
errShare
errSave
researcher View more