返回
Co-attention dictionary network for weakly-supervised semantic segmentation
DOI:10.1016/j.neucom.2021.11.046.png)
摘要
En 中文
In this paper, we propose the co-attention dictionary network (CODNet) for weakly-supervised semantic segmentation using only image-level class labels. The CODNet model exploits extra semantic information by jointly leveraging a pair of samples with common semantics through co-attention rather than processing them independently. The inter-sample similarities of spatially distributed deep features are computed to merge reference features through non-local connections. To discover similar patterns regardless of appearance variations, we propose to extract image representations by equipping the neural networks with dictionary learning which provides the universal basis elements for different images. Based on the CODNet model, we propose a multi-reference class activation map (MR-CAM) algorithm which generates semantic segmentation masks for a target image by jointly merging semantic cues from multiple reference images. Experimental results on the PASCAL VOC 2012 and MSCOCO benchmark data sets for weakly-supervised semantic segmentation show that the proposed algorithm performs favorably against the state-of-the-art methods.(c) 2021 Elsevier B.V. All rights reserved.
Keyword:
Weakly-supervised semantic segmentation
Dictionary learning
Co-attention
期刊
IF:
6.5
论文数:
2.5W
被引数:
6.5W
机构
引用论文
Semantic segmentation using stride spatial pyramid pooling and dual attention decoder
PATTERN RECOGNITION
IF7.6
Bilateral attention decoder: A lightweight decoder for real-time semantic segmentation
NEURAL NETWORKS
IF6.3
没有更多内容

