arrow
返回

A novel weight initialization with adaptive hyper-parameters for deep semantic segmentation

delete2021-03-20
delete2
PRE
AI
H
Haq, Nuhman Ui
A
Ahmad Khan
Z
Zia ur Rehman *
A
Ahmad Din
Ling Shao 封面图
Ling Shao (Ling Shao)
S
Sajid Shah
DOI:10.1007/s11042-021-10510-1delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
The semantic segmentation process divides an image into its constituent objects and background by assigning a corresponding class label to each pixel in the image. Semantic segmentation is an important area in computer vision with wide practical applications. The contemporary semantic segmentation approaches are primarily based on two types of deep neural networks architectures i.e., symmetric and asymmetric networks. Both types of networks consist of several layers of neurons which are arranged in two sections called encoder and decoder. The encoder section receives the input image and the decoder section outputs the segmented image. However, both sections in symmetric networks have the same number of layers and the number of neurons in an encoder layer is the same as that of the corresponding layer in the decoder section but asymmetric networks do not strictly follow such one-one correspondence between encoder and decoder layers. At the moment, SegNet and ESNet are the two leading state-of-the-art symmetric encoder-decoder deep neural network architectures. However, both architectures require extensive training for good generalization and need several hundred epochs for convergence. This paper aims to improve the convergence and enhance network generalization by introducing two novelties into the network training process. The first novelty is a weight initialization method and the second contribution is an adaptive mechanism for dynamic layer learning rate adjustment in training loop. The proposed initialization technique uses transfer learning to initialize the encoder section of the network, but for initialization of decoder section, the weights of the encoder section layers are copied to the corresponding layers of the decoder section. The second contribution of the paper is an adaptive layer learning rate method, wherein the learning rates of the encoder layers are updated based on a metric representing the difference between the probability distributions of the input images and encoder weights. Likewise, the learning rates of the decoder layers are updated based on the difference between the probability distributions of the output labels and decoder weights. Intensive empirical validation of the proposed approach shows significant improvement in terms of faster convergence and generalization.
Keyword:
Semantic segmentation
Deep learning
Initialization
Adaptive layer learning rate
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Multimedia Tools and Applications 封面图
Multimedia Tools and Applications
IF:
3
论文数:
2.0W
被引数:
3.2W

机构

C
comsats university islamabad (cui)
学者数:
1.1W
论文数: 1.1W
被引数: 7
引用论文

引用论文

Solution structure of the Lewis x oligosaccharide determined by NMR spectroscopy and molecular dynamics simulations
err2002-05-01
err0
PREAI
errKristine E. Miller; Chaitali Mukhopadhyay; Perseveranda Cagas; C. Allen Bush
err分享
err收藏
Validation of the Soil Moisture Active Passive (SMAP) satellite soil moisture retrieval in an Arctic tundra environment
err2017-05-14
err0
errOAAI
errElizabeth Wrona; Tracy L. Rowlandson; Manoj Nambiar; Aaron A. Berg; Andreas Colliander; Philip Marsh
err分享
err收藏
Determinants of quality of life in patients with fibromyalgia: A structural equation modeling approach纤维肌痛患者生活质量的决定因素: 结构方程建模方法
err2017-02-03
err0
errOAAI
errJeong-Won Lee; Kyung-Eun Lee; Dong-Jin Park; Seong-Ho Kim; Seong-Su Nah; Ji Hyun Lee; Seong-Kyu Kim; Yeon-Ah Lee; Seung-Jae Hong; Hyun-Sook Kim; Hye-Soon Lee; Hyoun Ah Kim; Chung-Il Joung; Sang-Hyon Kim; Shin-Seok Lee
err分享
err收藏
学者 查看更多内容