arrow
Return

3D shape classification based on global and local features extraction with collaborative learning

delete2023-09-28
delete1
PRE
AI
B
Bo Ding
张立保 (Libao Zhang)
何勇军 (Yongjun He) *
J
Jian Qin
DOI:10.1007/s00371-023-03098-0delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
It is important to extract both global and local features for view-based 3D shape classification. Therefore, we propose a 3D shape classification method based on global and local features extraction with collaborative learning. This method consists of a patch-level transformer sub-network (PTS) and a view-level transformer sub-network (VTS). In the PTS, a single view is divided into multiple patches. And a multi-layer transformer encoder is employed to accurately highlight discriminative patches and capture correlations among patches in a view, which can efficiently filter out the meaningless information and enhance meaningful information. The PTS can aggregate patch features into a 3D shape representation with rich local details. In the VTS, a multi-layer transformer encoder is employed to assign different attention to each view and obtain the contextual relationship among views, which can highlight the discriminative views among all the views of the same 3D shape and efficiently aggregate view features into a 3D shape representation. A collaborative loss is applied to encourage the two branches to learn collaboratively and teach each other in training. Experiments on two 3D benchmark datasets show that our proposed method outperforms current methods.
Keywords:
3D shape classification
Transformer encoder
Collaborative learning
Local features
Global features

Journal

Visual Computer cover
Visual Computer
IF:
2.9
Papers:
4.5K
Citations:
6.5K

Organization

H
harbin institute of technology
Scholars:
8.0W
Papers: 6.6W
Citations: 66