arrow
Return

Panoptic-VSNet: Visual-semantic prior knowledge-driven multimodal 3D panoptic segmentation

delete2026-02-05
delete0
PRE
AI
X
Xiao Li
李辉 cover
李辉 (Hui Li)
X
Xiangzhen Kong
Y
Yuang Ji
Z
Zhiyu Liu
H
Hao Liu
DOI:10.1016/j.patcog.2026.113239delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
• Maps points to semantic regions via CLIP-driven activation to reduce semantic gaps. • Fuses image panoptic masks and text embeddings for precise instance boundaries. • Employs dynamic kernels and multi-scale fusion for context capture and detail boost.
Keywords:
Panoptic segmentation
Visual-semantic prior
CLIP-driven activation
Multi-scale fusion
Dynamic kernels

Journal

Pattern Recognition cover
Pattern Recognition
IF:
7.6
Papers:
1.3W
Citations:
4.5W

Organization

E
eindhoven university of technology
Scholars:
985
Papers: 433
Citations: 0
O
ocean university of china
Scholars:
3.1W
Papers: 2.0W
Citations: 21
Q
qingdao turing technology co., ltd
Scholars:
1
Papers: 1
Citations: 0
Q
Qingdao University of Science and Technology
Scholars:
1.5K
Papers: 441
Citations: 2.2W
researcher View more organizations