arrow
返回

MVFormer: Diversifying feature normalization and token mixing for efficient vision transformers

delete2025-07-23
delete0
PRE
AI
J
Jongseong Bae
S
Susang Kim
M
Minsu Cho
H
Ha Young Kim *
DOI:10.1016/j.patrec.2025.07.019delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
• 我们提出MVFormer通过token混合器和归一化进行多样化特征学习。• MVN结合了三种类型的归一化,反映多样化的特征分布。• MVTM通过每阶段多样化感受野实现阶段特异性。• 同时采用MVN和MVTM增强了多样化视角的能力。• MVFormer在ImageNet-1K基准测试中超越了现有的基于卷积的ViT。
Keyword:
MVFormer
token mixer
normalization
receptive field
diverse feature learning

期刊

Pattern Recognition Letters 封面图
Pattern Recognition Letters
IF:
3.3
论文数:
8.0K
被引数:
1.6W

机构

P
POSTECH
学者数:
936
论文数: 383
被引数: 7
Y
Yonsei University
学者数:
4.8W
论文数: 4.6W
被引数: 5.2W
引用论文

引用论文

暂无论文信息