arrow
Return

High-resolution rectified gradient-based visual explanations for weakly supervised segmentation

delete2022-09-01
delete4
PRE
AI
T
Tianyou Zheng *
王强 (Qiang Wang)
Y
Yue Shen
马祥 (Xiang Ma)
X
Xiaotian Lin
DOI:10.1016/j.patcog.2022.108724delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Visual explanations for convolutional neural networks (CNNs) act as the backbone for weakly supervised segmentation with image-level labels. This paper proposes a high-resolution rectified gradient-based class activation mapping with bounding box annotations (bbox) to improve the initial seed for weakly supervised segmentation (WSS) tasks. HRCAM extends Grad-CAM by separating the gradient maps from the class activation maps from the shallow layer for higher resolution. Gradient rectified methods are proposed to improve the visualization and WSS score. Experiments and evaluations are conducted to verify the performance of HRCAM-BB on Pascal VOC 2012 and COCO datasets. On Pascal VOC 2012 set, our method achieves outstanding performance with a mean intersection over union (mIOU) of 71.6 with image-level labels and 78.2 with bbox on WSSS, and increases the WSIS mIOU (AP(50)) to 52.1 with image-level labels, and 61.9 with bbox. our method surpasses the previous SOTA approach in the same condition. (C) 2022 Elsevier Ltd. All rights reserved.

Journal

Pattern Recognition cover
Pattern Recognition
IF:
7.6
Papers:
1.3W
Citations:
4.5W

Organization

H
harbin institute of technology
Scholars:
8.0W
Papers: 6.6W
Citations: 66