arrow
返回

Lightweight transformer image feature extraction network

delete2024-01-31
delete84
delete
OA
AI
W
Wenfeng Zheng
S
Siyu Lu
Y
Youshuai Yang
Z
Zhengtong Yin
L
Lirong Yin *
DOI:10.7717/peerj-cs.1755delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
In recent years, the image feature extraction method based on Transformer has become a research hotspot. However, when using Transformer for image feature extraction, the model's complexity increases quadratically with the number of tokens entered. The quadratic complexity prevents vision transformer-based backbone networks from modelling high-resolution images and is computationally expensive. To address this issue, this study proposes two approaches to speed up Transformer models. Firstly, the self-attention mechanism's quadratic complexity is reduced to linear, enhancing the model's internal processing speed. Next, a parameter-less lightweight pruning method is introduced, which adaptively samples input images to filter out unimportant tokens, effectively reducing irrelevant input. Finally, these two methods are combined to create an efficient attention mechanism. Experimental results demonstrate that the combined methods can reduce the computation of the original Transformer model by 30%-50%, while the efficient attention mechanism achieves an impressive 60%-70% reduction in computation.
Keyword:
Transformer
Quadratic complexity
Image feature extraction
Self-attention mechanism
Pruning
Efficient attention
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

PeerJ Computer Science 封面图
PeerJ Computer Science
IF:
2.5
论文数:
3.4K
被引数:
6.9K

机构

L
louisiana state university system
学者数:
2.3W
论文数: 2.0W
被引数: 15
G
guizhou university
学者数:
2.5W
论文数: 1.3W
被引数: 15