arrow
返回

A real-time semantic segmentation model using iteratively shared features in multiple sub-encoders

delete2023-08-01
delete22
PRE
AI
T
Tanmay Singha *
D
Duc-Son Pham
A
Aneesh Krishna
DOI:10.1016/j.patcog.2023.109557delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Recent studies show a significant growth in semantic segmentation. However, many semantic segmen-tation models still have a large number of parameters, making them unsuitable for resource-constrained embedded devices. To address this issue, we propose an efficient Shared Feature Reuse Segmentation (SFRSeg) model containing several novelties: a new yet effective shared-branch multiple sub-encoders design, a context mining module and a semantic aggregating module for better context granularity. In particular, our shared-branch approach improves the entire feature hierarchy by sharing the spatial and context knowledge in both shallow and deep branches. After every shared point in each sub-encoder, a proposed cascading context mining (CCM) module is deployed to filter out the noisy spatial details from the feature maps and provides a diverse size of receptive fields for capturing the latent context between multi-scale geometric shapes in the scene. To overcome the gradient vanishing issue at the early stage, we reduce the number of layers in the first sub-encoder and employ a unique multiple sub-encoders design which reprocesses the rich global feature maps through multiple sub-encoders for better feature refinement. Later, the rich semantic features generated by the efficient sub-encoders at different levels are fused by the proposed Hybrid Path Attention Semantic Aggregation (HPA-SA) module that effectively reduces the semantic gap between feature maps at different levels and alleviate the well-known bound-ary degeneration effect. To make it computationally efficient for resource-constrained embedded devices, a series of lightweight methods such as a lightweight encoder, a squeeze-and-excitation design, separa-ble convolution filters, channel reduction (CR) are carefully exploited. With an exceptional performance on Cityscapes (70.6% test mIoU) and CamVid (74.7% test mIoU) data sets, the proposed model is shown to be superior over existing light real-time semantic segmentation models whilst having only 1.6 million parameters. (c) 2023 Elsevier Ltd. All rights reserved.
Keyword:
Semantic segmentation
Deep convolution neural networks
Multi-encoder
Decoder
Feature scaling
Feature aggregation
Feature reuse
Resource-constrained applications
Mobile devices

期刊

Pattern Recognition 封面图
Pattern Recognition
IF:
7.6
论文数:
1.3W
被引数:
4.5W

机构

C
Curtin University
学者数:
1.5W
论文数: 1.8W
被引数: 2.8W
引用论文

引用论文

Electrochemically assisted micro localized grafting of aptamers in a microchannel engraved in fluorinated thermoplastic polymer Dyneon THV
err2015-01-01
err0
PREAI
errC. Perréard; Y. Ladner; F. d'Orlyé; S. Descroix; V. Taniga; A. Varenne; F. Kanoufi; C. Slim; S. Griveau; F. Bedioui
err分享
err收藏
Deep gated attention networks for large-scale street-level scene segmentation
err2019-04-01
err81
PREAI
errZhang, Pingping; Liu, Wei; Wang, Hongyu; Lei, Yinjie; Lu, Huchuan
err分享
err收藏
Video semantic segmentation via feature propagation with holistic attention
err2020-08-01
err24
PREAI
errWu, Junrong; Wen, Zongzheng; Zhao, Sanyuan; Huang, Kele
err分享
err收藏
Contextual ensemble network for semantic segmentation
err2022-02-01
err65
errOAAI
errZhou, Quan; Wu, Xiaofu; Zhang, Suofei; Kang, Bin; Ge, Zongyuan; Latecki, Longin Jan
err分享
err收藏
Semantic object classes in video: A high-definition ground truth database
err2009-01-01
err1.1K
PREAI
errBrostow, Gabriel J.; Fauqueur, Julien; Cipolla, Roberto
err分享
err收藏
Efficient semantic segmentation with pyramidal fusion
err2021-02-01
err73
PREAI
errOrsic, Marin; Segvic, Sinisa
err分享
err收藏
Contextual deconvolution network for semantic segmentation用于语义分割的上下文反卷积网络
err2020-05-01
err46
PREAI
errFu, Jun; Liu, Jing; Li, Yong; Bao, Yongjun; Yan, Weipeng; Fang, Zhiwei; Lu, Hanqing
err分享
err收藏
学者 查看更多内容