arrow
返回

STAN: A sequential transformation attention-based network for scene text recognition

delete2021-03-01
delete38
PRE
AI
Q
Qingxiang Lin
C
Canjie Luo
金
金连文 (Lianwen Jin)
S
Songxuan Lai *
DOI:10.1016/j.patcog.2020.107692delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Scene text with an irregular layout is difficult to recognize. To this end, a Sequential Transformation Attention-based Network (STAN), which comprises a sequential transformation network and an attention-based recognition network, is proposed for general scene text recognition. The sequential transformation network rectifies irregular text by decomposing the task into a series of patch-wise basic transformations, followed by a grid projection submodule to smooth the junction between neighboring patches. The entire rectification process is able to be trained in an end-to-end weakly supervised manner, requiring only images and their corresponding groundtruth text. Based on the rectified images, an attention-based recognition network is employed to predict a character sequence. Experiments on several benchmarks demonstrate the state-of-the-art performance of STAN on both regular and irregular text. (C) 2020 Elsevier Ltd. All rights reserved.
Keyword:
Scene text recognition
Scene text rectification
Optical character recognition
Deep learning
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Pattern Recognition 封面图
Pattern Recognition
IF:
7.6
论文数:
1.3W
被引数:
4.5W

机构

S
south china university of technology
学者数:
6.8W
论文数: 5.1W
被引数: 85
引用论文

引用论文

Recent advances in convolutional neural networks卷积神经网络的最新进展
err2018-05-01
err3.8K
errOAAI
errGu, Jiuxiang; Wang, Zhenhua; Kuen, Jason; Ma, Lianyang; Shahroudy, Amir; Shuai, Bing; Liu, Ting; Wang, Xingxing; Wang, Gang; Cai, Jianfei; Chen, Tsuhan
err分享
err收藏
err分享
err收藏
Integration of RF-MEMS resonators on submicrometric commercial CMOS technologies
err2008-11-27
err0
PREAI
errJ L Lopez; J Verd; J Teva; G Murillo; J Giner; F Torres; A Uranga; G Abadal; N Barniol
err分享
err收藏
Could scene context be beneficial for scene text detection?
err2016-10-01
err34
PREAI
errZhu, Anna; Gao, Renwu; Uchida, Seiichi
err分享
err收藏
A blind deconvolution model for scene text detection and recognition in video
err2016-06-01
err37
PREAI
errKhare, Vijeta; Shivakumara, Palaiahnakote; Raveendran, Paramesran; Blumenstein, Michael
err分享
err收藏
Curved scene text detection via transverse and longitudinal sequence connection
err2019-06-01
err187
errOAAI
errLiu, Yuliang; Jin, Lianwen; Zhang, Shuaitao; Luo, Canjie; Zhang, Sheng
err分享
err收藏
End-to-end scene text recognition using tree-structured models
err2014-09-01
err27
PREAI
errShi, Cunzhao; Wang, Chunheng; Xiao, Baihua; Gao, Song; Hu, Jinlong
err分享
err收藏
学者 查看更多内容