返回
Tensor representation learning based image patch analysis for text identification and recognition
DOI:10.1016/j.patcog.2014.09.025.png)
摘要
En 中文
In this paper, we introduce a novel framework for text identification and recognition, called tensor representation learning based image patch analysis (TRL-IPA). Unlike most of previous text identification approaches, which can only be applied to binarized images, TRL-IPA can be directly applied to gray level and color images. TRL-IPA is built on a general formulation of the convergent tensor representation learning (CTRL) algorithms. In the implementation of TRL-IPA, image patches are represented in the form of tensors, while low dimensional representations of these tensors are learned via a CTRL algorithm. To identify text regions in new coming document images, a random forest classifier is trained in the learned tensor subspace. Moreover, the TRL-IPA framework can be straightforwardly applied to recognition problems, such as handwritten digits recognition. We conducted extensive experiments on ancient Chinese, Arabic and Cyrillic document images, to evaluate TRL-IPA on text identification tasks. Experimental results demonstrate its effectiveness and robustness. In addition, recognition results on images of handwritten digits show its advantage over state-of-the-art vector and tensor representation based approaches. (C) 2014 Elsevier Ltd. All rights reserved.
Keyword:
Tensor representation learning
Convergence
Ancient document understanding
Text identification
Text recognition
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
IF:
7.6
论文数:
1.3W
被引数:
4.5W
机构
引用论文
A simplified approach to the HMM based texture analysis and its application to document segmentation
Forty years of research in character and document recognition-an industrial perspective
PATTERN RECOGNITION
IF7.6
Identification and Characterisation CRN Effectors in Phytophthora capsici Shows Modularity and Functional Diversity
PLoS ONE
IF0

