arrow
Return

EFRNet: Efficient Feature Reconstructing Network for Real-Time Scene Parsing

delete2022-01-01
delete2
PRE
AI
X
Xin Li
F
Fan Yang *
A
Ao Luo
Z
Zhicheng Jiao
程洪 (Hong Cheng)
Z
Zicheng Liu
DOI:10.1109/TMM.2021.3089422delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
In this paper, we introduce a light-weight and powerful convolutional neural network, termed as efficient feature reconstructing network (EFRNet), for real-time scene parsing. Our key idea is to decompose the process of learning high-resolution representations into two stages: i) bottom-up codebook/coding matrix learning and ii) top-down feature reconstructing. Specifically, the bottom-up process focuses on learning image-specific codewords (codebook) using deep-layer features and generating a coding matrix with the shallow-layer feature map. In the top-down process, the learned codebook and coding matrix are used to rebuild high-resolution features via a lightweight feature reconstructing operator (FRO). In addition, our EFRNet is constructed on a new building block, named efficient adaptive abstraction (EAA) block, to further reduce the overall network parameters and achieve a significant speed up. Extensive experiments are conducted on challenging benchmarks, such as CamVid and Cityscapes. The results show that EFRNet demonstrates state-of-the-art performance with an optimal balance between accuracy and speed.
Keywords:
Image reconstruction
Semantics
Encoding
Feature extraction
Convolutional codes
Real-time systems
Task analysis
Scene parsing
dictionary learning
representation learning
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

IEEE Transactions on Multimedia cover
IEEE Transactions on Multimedia
IF:
9.7
Papers:
4.5K
Citations:
2.4W

Organization

U
university of pennsylvania
Scholars:
9.2W
Papers: 7.8W
Citations: 153