arrow
返回

Dynamic attention guider network

delete2024-07-30
delete0
PRE
AI
C
Chunguang Yue
李
李金宝 (Jinbao Li) *
Q
Qichen Wang
D
Donghuan Zhang
DOI:10.1007/s00607-024-01328-4delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Hybrid networks, benefiting from both CNNs and Transformers architectures, exhibit stronger feature extraction capabilities compared to standalone CNNs or Transformers. However, in hybrid networks, the lack of attention in CNNs or insufficient refinement in attention mechanisms hinder the highlighting of target regions. Additionally, the computational cost of self-attention in Transformers poses a challenge to further improving network performance. To address these issues, we propose a novel point-to-point Dynamic Attention Guider(DAG) that dynamically generates multi-scale large receptive field attention to guide CNN networks to focus on target regions. Building upon DAG, we introduce a new hybrid network called the Dynamic Attention Guider Network (DAGN), which effectively combines Dynamic Attention Guider Block (DAGB) modules with Transformers to alleviate the computational cost of self-attention in processing high-resolution input images. Extensive experiments demonstrate that the proposed network outperforms existing state-of-the-art models across various downstream tasks. Specifically, the network achieves a Top-1 classification accuracy of 88.3% on ImageNet1k. For object detection and instance segmentation on COCO, it respectively surpasses the best FocalNet-T model by 1.6 APb\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$AP<^>b$$\end{document} and 1.5 APm\documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$AP<^>m$$\end{document}, while achieving a top performance of 48.2% in semantic segmentation on ADE20K.
Keyword:
Hybrid networks
Multi-scale
Multi-path
Attention
Backbone

期刊

C
Computing
IF:
2.8
论文数:
2.3K
被引数:
3.5K

机构

Q
Qilu University of Technology
学者数:
1.1W
论文数: 8.9K
被引数: 16
H
Heilongjiang University
学者数:
8.5K
论文数: 5.2K
被引数: 6.8K
引用论文

引用论文

A Pilot Study of the Efficacy of the Unified Protocol for Transdiagnostic Treatment of Emotional Disorders in Treating Posttraumatic Psychopathology: A Randomized Controlled Trial
err2021-01-16
err0
errOAAI
errMeaghan L. O'Donnell; Winnie Lau; Katherine Chisholm; James Agathos; Jonathon Little; Sonia Terhaag; Rachel Brand; Andrea Putica; Alexander C. N. Holmes; Lynda Katona; Kim L. Felmingham; Kim Murray; Fardous Hosseiny; Matthew W. Gallagher
err分享
err收藏
Factors and conditions for the development of the digital economy in Russia
err2021-03-19
err0
errOAAI
errElvira Karieva; Liliya Akhmetshina; Olga Fokina
err分享
err收藏
err分享
err收藏
The role of visual short-term memory in empty cell localization
err2005-11-01
err0
errOAAI
errAndrew Hollingworth; Joo-Seok Hyun; Weiwei Zhang
err分享
err收藏
学者 查看更多内容