arrow
返回

Semantic-Aware Message Broadcasting for Efficient Unsupervised Domain Adaptation

delete2024-01-01
delete0
delete
OA
AI
X
Xin Li
C
Cuiling Lan *
G
Guoqiang Wei
陈志波 封面图
陈志波 (Zhibo Chen) *
DOI:10.1109/TIP.2024.3437212delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Vision transformer has demonstrated great potential in abundant vision tasks. However, it also inevitably suffers from poor generalization capability when the distribution shift occurs in testing (i.e., out-of-distribution data). To mitigate this issue, we propose a novel method, Semantic-aware Message Broadcasting (SAMB), which enables more informative and flexible feature alignment for unsupervised domain adaptation (UDA). Particularly, we study the attention module in the vision transformer and notice that the alignment space using one global class token lacks enough flexibility, where it interacts information with all image tokens in the same manner but ignores the rich semantics of different regions. In this paper, we aim to improve the richness of the alignment features by enabling semantic-aware adaptive message broadcasting. Particularly, we introduce a group of learned group tokens as nodes to aggregate the global information from all image tokens, but encourage different group tokens to adaptively focus on the message broadcasting to different semantic regions. In this way, our message broadcasting encourages the group tokens to learn more informative and diverse information for effective domain alignment. Moreover, we systematically study the effects of adversarial-based feature alignment (ADA) and pseudo-label based self-training (PST) on UDA. We find that one simple two-stage training strategy with the cooperation of ADA and PST can further improve the adaptation capability of the vision transformer. Extensive experiments on DomainNet, OfficeHome, and VisDA-2017 demonstrate the effectiveness of our methods for UDA.
Keyword:
Transformers
Broadcasting
Training
Computer vision
Task analysis
Semantics
Adaptation models
Unsupervised domain adaptation
vision transformer
semantic-aware message broadcasting
adversarial-based feature alignment
self-training

期刊

IEEE Transactions on Image Processing 封面图
IEEE Transactions on Image Processing
IF:
13.7
论文数:
1.0W
被引数:
8.4W

机构

U
university of science & technology of china, cas
学者数:
3.2W
论文数: 2.7W
被引数: 74
C
chinese academy of sciences
学者数:
56.7W
论文数: 45.0W
被引数: 704
引用论文

引用论文

Perfectly Prep
err
IF0
err2008-06-26
err0
PREAI
errSarah A. Chase
err分享
err收藏
err分享
err收藏
Using SEPIC Topology for Improving Power Factor in Distributed Power Supply Systems
err2015-09-22
err0
PREAI
errJ. Sebastián; J. Uceda; J.A. Cobos; J. Arau
err分享
err收藏
The mitochondrial complex I inhibitor rotenone triggers a cerebral tauopathy
err2005-10-10
err0
PREAI
errGünter U. Höglinger; Annie Lannuzel; Myriam Escobar Khondiker; Patrick P. Michel; Charles Duyckaerts; Jean Féger; Pierre Champy; Annick Prigent; Fadia Medja; Anne Lombes; Wolfgang H. Oertel; Merle Ruberg; Etienne C. Hirsch
err分享
err收藏
学者 查看更多内容