arrow
Return

Learning multi-view visual correspondences with self-supervision

delete2022-04-01
delete15
PRE
AI
P
Pengcheng Zhang
L
Lei Zhou
X
Xiao Bai *
C
Chen Wang
周军 (Jun Zhou)
L
Liang Zhang
J
Jin Zheng
DOI:10.1016/j.displa.2022.102160delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Stereo-based 3D reconstruction requires to match features across images captured from slightly different viewing angles to recover 3D coordinates of the image pixels. Despite the workload of collecting data, annotating matched pixels requires also heavy labor. As recent researches for self-supervised representation learning has gained great progress, learning multi-view visual correspondences from large scale raw videos serves as an alternative. However, existing methods which benefit from contrastive learning tend to neglect false negative samples when matching between adjacent frames in a video, leading to sub-optimal optimization for visual features. In this paper, we propose a contrastive learning framework that construct self-supervision by semi-global visual correspondence to alleviate learning degradation when false negatives are involved in training. Our learning framework consists of pixel-level contrastive learning via patch reconstruction and patch-level contrastive learning cross videos. We also introduce saliency guidance to extract salient regions from video frames to further reduce potential false negatives. By optimizing the model with the proposed semi-global contrastive learning method, learned representations are forced to be discriminative and robust. Experiments demonstrate that our proposed method outperforms previous self-supervised methods on video object segmentation tasks. Moreover, when compared to fully-supervised algorithms designed for specific tasks, our proposed method also achieves competitive results.
Keywords:
Visual correspondence
Self-supervised learning
Contrastive learning

Journal

Displays cover
Displays
IF:
3.4
Papers:
2.1K
Citations:
3.2K

Organization

B
Beihang University
Scholars:
5.1W
Papers: 4.1W
Citations: 37
G
Griffith University
Scholars:
1.5W
Papers: 1.6W
Citations: 2.5W