arrow
Return

Semantic Segmentation Network combining Gaussian Perception and Iterative Multi-Scale Attention

delete2025-08-21
delete0
PRE
AI
Z
Zunwang Ke
王国盛 cover
王国盛 (Guosheng Wang)
Y
Yugui Zhang *
Y
Yunlong Shi
F
Fengyu Guo
Y
Yuelin Zou
Z
Zhaofan Li
R
Run Guo
Z
Zhou Ji-sheng
DOI:10.1007/s00530-025-01923-1delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Semantic segmentation is crucial in autonomous driving, offering exceptional scene understanding to tackle challenges like small object edges and blurred textures in complex traffic environments. By performing pixel-level classification, it provides vehicles with comprehensive environmental information, ensuring safe navigation. However, when dealing with fine, fuzzy-bordered objects in complex scenes, the existing techniques face the issue of low segmentation accuracy due to their insufficient feature extraction capability. To address this problem, this study proposes a semantic segmentation network that combines Gaussian perception with iterative multi-scale attention. The method integrates Gaussian perception with local–global channel attention, accurately models pixel associations, and dynamically focuses on features to address the issue of edge blurring in complex scene segmentation. At the same time, the method employs the difference module to enhance the low-frequency features, and integrates the iterative multi-scale attention mechanism to achieve deep integration of low-frequency and high-frequency information. This enhances the fine capture of features and mitigates the edge discontinuity issue in segmentation caused by the masking of boundary information. In addition, the method combines channel and spatial attention to optimize feature extraction, enhance the sensory field, and improve detail, context, and boundary recognition abilities. This significantly improves the feature expression ability and reduces the probability of background mis-segmentation. The experimental results show that the proposed method achieves 79.34% mIoU on the Cityscapes validation set (a 1.32% improvement over PIDNet-S) and 81.48% mIoU on the CamVid test set (a 1.05% improvement over PIDNet-S). These results demonstrate significant improvements over existing state-of-the-art methods, confirming the effectiveness of this approach in semantic segmentation of complex urban scenes. The source code has been made publicly available on GitHub: https://github.com/wgsheng897/GMSANet.git
Keywords:
Semantic segmentation
Gaussian function
Local–global features
Spatial and channel attention mechanisms

Journal

Multimedia Systems cover
Multimedia Systems
IF:
3.1
Papers:
2.7K
Citations:
2.7K

Organization

C
College of Computer and Information Engineering
Scholars:
100
Papers: 38
Citations: 0
T
telecommunication company
Scholars:
6
Papers: 2
Citations: 0
P
Peking University Third Hospital
Scholars:
1.4K
Papers: 411
Citations: 6.8K
I
Institute of Semiconductors
Scholars:
369
Papers: 133
Citations: 5.7K
F
first medical center
Scholars:
55
Papers: 18
Citations: 0
S
School of Software
Scholars:
360
Papers: 144
Citations: 0
D
Department of Dermatology
Scholars:
2.4K
Papers: 909
Citations: 8
researcher View more organizations