arrow
Return

Object representation enhancement for self-supervised colocalization

delete2022-06-13
delete1
delete
OA
AI
H
Huifang Li
李浥东 (Yidong Li) *
金一 (Yi Jin)
T
Tao Wang
DOI:10.1002/int.22938delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Self-supervised colocalization is to localize common objects in the data set containing only one superclass without using human-annotated labels. Existing methods achieve impressive results by employing self-supervised pretext learning. However, a common limitation still exists. They either tend to overextend activations to the background, or they tend to activate the most discriminative object part. To alleviate this problem, we propose an object representation enhancement model to weaken background distraction and to mine complementary object regions during the object representation learning. Specifically, we first propose an Object-aware Representation Enhancement (ORE) module to estimate an object mask for each input image, guiding the model to disregard the background content and focus on the foreground object. The ORE module and the subsequent self-supervised learning can mutually reinforce each other. Then we propose a Masked Self-supervised Learning branch and design a masked attention consistency objective to induce the model to activate complementary parts of the object effectively. Extensive experiments on four fine-grained data sets demonstrate the superiority of the proposed model.
Keywords:
colocalization
contrastive self-supervised learning
masked self-supervised learning

Journal

International Journal of Intelligent Systems cover
International Journal of Intelligent Systems
IF:
3.7
Papers:
3.0K
Citations:
8.1K

Organization

B
Beijing Jiaotong University
Scholars:
2.2W
Papers: 1.7W
Citations: 1.2W