arrow
返回

End-to-end scene text recognition using tree-structured models

delete2014-09-01
delete27
PRE
AI
C
Cunzhao Shi
C
Chunheng Wang *
B
Baihua Xiao
S
Song Gao
胡锦龙 封面图
胡锦龙 (Jinlong Hu)
DOI:10.1016/j.patcog.2014.03.023delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Detecting and recognizing text in natural images are quite challenging and have received much attention from the computer vision community in recent years. In this paper, we propose a robust end-to-end scene text recognition method, which utilizes tree-structured character models and normalized pictorial structured word models. For each category of characters, we build a part-based tree-structured model (TSM) so as to make use of the character-specific structure information as well as the local appearance information. The TSM could detect each part of the character and recognize the unique structure as well, seamlessly combining character detection and recognition together. As the TSMs could accurately detect characters from complex background, for text localization, we apply TSMs for all the characters on the coarse text detection regions to eliminate the false positives and search the possible missing characters as well. While for word recognition, we propose a normalized pictorial structure (PS) framework to deal with the bias caused by words of different lengths. Experimental results on a range of challenging public datasets (ICDAR 2003, ICDAR 2011, SVT) demonstrate that the proposed method outperforms state-of-the-art methods both for text localization and word recognition. (C) 2014 Elsevier Ltd. All rights reserved.
Keyword:
End-to-end
Scene text recognition
Part-based tree-structured models (TSMs)
Normalized pictorial structure
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Pattern Recognition 封面图
Pattern Recognition
IF:
7.6
论文数:
1.3W
被引数:
4.5W

机构

C
chinese academy of sciences
学者数:
56.7W
论文数: 45.0W
被引数: 704
引用论文

引用论文

err分享
err收藏
Comparative analysis of occlusion methods for artificial sphincters人工括约肌闭塞方法的比较分析
err2020-04-07
err0
PREAI
errLeonardo Marziale; Gioia Lucarini; Tommaso Mazzocchi; Leonardo Ricotti; Arianna Menciassi
err分享
err收藏
Cybercrime and virtual offender convergence settings
err2012-05-25
err0
PREAI
errMelvin R. J. Soudijn; Birgit C. H. T Zegers
err分享
err收藏
err分享
err收藏
Conditional random field for text segmentation from images with complex background
err2010-10-01
err16
PREAI
errLi, Minhua; Bai, Meng; Wang, Chunheng; Xiao, Baihua
err分享
err收藏
err分享
err收藏
学者 查看更多内容