arrow
返回

MGRW-Transformer: Multigranularity Random Walk Transformer Model for Interpretable Learning

delete2025-01-01
delete2
PRE
AI
丁卫平 封面图
丁卫平 (Weiping Ding) *
Y
Yu Geng
J
Jiashuang Huang
H
Hengrong Ju
H
Haipeng Wang
C
Chin‐Teng Lin
DOI:10.1109/TNNLS.2023.3326283delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Deep-learning models have been widely used in image recognition tasks due to their strong feature-learning ability. However, most of the current deep-learning models are black box systems that lack a semantic explanation of how they reached their conclusions. This makes it difficult to apply these methods to complex medical image recognition tasks. The vision transformer (ViT) model is the most commonly used deep-learning model with a self-attention mechanism that shows the region of influence as compared to traditional convolutional networks. Thus, ViT offers greater interpretability. However, medical images often contain lesions of variable size in different locations, which makes it difficult for a deep-learning model with a self-attention module to reach correct and explainable conclusions. We propose a multigranularity random walk transformer (MGRW-Transformer) model guided by an attention mechanism to find the regions that influence the recognition task. Our method divides the image into multiple subimage blocks and transfers them to the ViT module for classification. Simultaneously, the attention matrix output from the multiattention layer is fused with the multigranularity random walk module. Within the multigranularity random walk module, the segmented image blocks are used as nodes to construct an undirected graph using the attention node as a starting node and guiding the coarse-grained random walk. We appropriately divide the coarse blocks into finer ones to manage the computational cost and combine the results based on the importance of the discovered features. The result is that the model offers a semantic interpretation of the input image, a visualization of the interpretation, and insight into how the decision was reached. Experimental results show that our method improves classification performance with medical images while presenting an understandable interpretation for use by medical professionals.
Keyword:
Graph random walk
interpretable method
multigranularity formal analysis
self-attention mechanism
vision transformer (ViT)

期刊

IEEE Transactions on Neural Networks and Learning Systems 封面图
IEEE Transactions on Neural Networks and Learning Systems
IF:
8.9
论文数:
7.5K
被引数:
7.2W

机构

U
university of technology sydney
学者数:
1.6W
论文数: 2.0W
被引数: 25
N
Nantong University
学者数:
1.9W
论文数: 1.1W
被引数: 2.0W
引用论文

引用论文

Lung cancer subtype classification using histopathological images based on weakly supervised multi-instance learning基于弱监督多示例学习的组织病理学图像肺癌亚型分类
err2021-12-02
err12
PREAI
errZhao, Lu; Xu, Xiaowei; Hou, Runping; Zhao, Wangyuan; Zhong, Hai; Teng, Haohua; Han, Yuchen; Fu, Xiaolong; Sun, Jianqi; Zhao, Jun
err分享
err收藏
Theoretical analysis of an impact-bistable piezoelectric energy harvester
err2019-05-06
err0
PREAI
errZhengqiu Xie; C. A. Kitio Kwuimy; Tao Wang; Xiaoxi Ding; Wenbin Huang
err分享
err收藏
5-axis double-flank CNC machining of spiral bevel gears via custom-shaped tools—Part II: physical validations and experiments
err2021-11-25
err0
errOAAI
errGaizka Gómez Escudero; Pengbo Bo; Haizea González-Barrio; Amaia Calleja-Ochoa; Michael Bartoň; Luis Norberto López de Lacalle
err分享
err收藏
ImageNet Large Scale Visual Recognition ChallengeImageNet大规模视觉识别挑战
err2015-04-11
err2.7W
PREAI
errRussakovsky, Olga; Deng, Jia; Su, Hao; Krause, Jonathan; Satheesh, Sanjeev; Ma, Sean; Huang, Zhiheng; Karpathy, Andrej; Khosla, Aditya; Bernstein, Michael; Berg, Alexander C.; Fei-Fei, Li
err分享
err收藏
err分享
err收藏
学者 查看更多内容