arrow
返回

Adaptive embedding gate for attention-based scene text recognition

delete2020-03-01
delete32
delete
OA
AI
X
Xiaoxue Chen
T
Tianwei Wang
Y
Yuanzhi Zhu
金
金连文 (Lianwen Jin) *
C
Canjie Luo
DOI:10.1016/j.neucom.2019.11.049delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
Scene text recognition has attracted particular research interest because it is a very challenging problem and has various applications. The most cutting-edge methods are attentional encoder-decoder frameworks that learn the alignment between the input image and output sequences. In particular, the decoder recurrently outputs predictions, using the prediction of the previous step as a guidance for every time step. In this study, we point out that the inappropriate use of previous predictions in existing attentional decoders restricts the recognition performance and brings instability. To handle this problem, we propose a novel module, namely adaptive embedding gate (AEG). The proposed AEG focuses on introducing high-order character language models to attentional decoders by controlling the information transmission between adjacent characters. AEG is a flexible module and can be easily integrated into the state-of-the-art attentional decoders for scene text recognition. We evaluate its effectiveness as well as robustness on a number of standard benchmarks, including the IIIT5K, SVT, SVT-P, CUTE80, and ICDAR datasets. Experimental results demonstrate that AEG can significantly boost recognition performance and bring better robustness. (C) 2019 The Author(s). Published by Elsevier B.V.
Keyword:
Deep learning
Scene text recognition
Attention mechanism
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Neurocomputing 封面图
Neurocomputing
IF:
6.5
论文数:
2.5W
被引数:
6.5W

机构

S
south china university of technology
学者数:
6.8W
论文数: 5.1W
被引数: 85
引用论文

引用论文

Reading scene text with fully convolutional sequence modeling
err2019-04-01
err51
PREAI
errGao, Yunze; Chen, Yingying; Wang, Jinqiao; Tang, Ming; Lu, Hanqing
err分享
err收藏
err分享
err收藏
Multi-oriented text detection and verification in video frames and scene images
err2018-01-01
err27
errOAAI
errSain, Aneeshan; Bhunia, Ayan Kumar; Roy, Partha Pratim; Pal, Umapada
err分享
err收藏
err分享
err收藏
学者 查看更多内容