Return
Robust visual place recognition with adaptive deformable token aggregation
DOI:10.1016/j.cag.2025.104342.png)
Abstract
En 中文
• Proposes ADTA-VPR, a robust framework for visual place recognition under complex conditions. • Introduces Deformable Token Merging to adaptively downsample features with spatial flexibility. • Designs Prototype-guided Token Aggregation to suppress redundancy and enhance distinctiveness. • Constructs FCBR, a VPR benchmark with realistic photographic degradations. • Achieves state-of-the-art performance on diverse and challenging VPR benchmarks.
Keywords:
ADTA-VPR
Deformable Token Merging
Prototype-guided Token Aggregation
FCBR benchmark
Visual Place Recognition
Journal
C
IF:
4.2
Papers:
1.4K
Citations:
3.3K

