返回
Prior tokenization-based interactive segmentation with Vision Transformers
DOI:10.1016/j.patcog.2025.112361.png)
摘要
En 中文
• 先验token在交互式信息中处理语义特征的效果优于距离图。
• 在判别性约束下,先验token更容易区分目标和背景的语义差异。
• 交叉注意力紧密地将图像块token与用户意图对齐。
• 寄存器方法有效解决了交互式分割中的伪影问题。
Keyword:
prior tokens
cross-attention
interactive segmentation
semantic features
discriminative constraints
期刊
IF:
7.6
论文数:
1.3W
被引数:
4.5W
机构
引用论文
A fully convolutional two-stream fusion network for interactive image segmentation
NEURAL NETWORKS
IF6.3
CSANet: Cross-self attention guided by semantic click embedding for interactive segmentationCSANet:基于语义点击嵌入引导的交叉自注意力机制用于交互式分割
Text-video retrieval re-ranking via multi-grained cross attention and frozen image encoders
PATTERN RECOGNITION
IF7.6

