Return
Prior tokenization-based interactive segmentation with Vision Transformers
DOI:10.1016/j.patcog.2025.112361.png)
Abstract
En 中文
• Prior tokens handle semantic features in interactive information better than distance maps. • Under discriminative constraints, prior tokens distinguish target and background semantic divergence more easily. • Cross-attention tightly aligns image patch tokens with user intent. • The register method effectively addresses artifacts in interactive segmentation.
Keywords:
prior tokens
cross-attention
interactive segmentation
semantic features
discriminative constraints
Journal
IF:
7.6
Papers:
1.3W
Citations:
4.5W
Organization
Cited Papers
A fully convolutional two-stream fusion network for interactive image segmentation
NEURAL NETWORKS
IF6.3
Text-video retrieval re-ranking via multi-grained cross attention and frozen image encoders
PATTERN RECOGNITION
IF7.6

