返回
Word level multi-script identification
DOI:10.1016/j.patrec.2008.01.027.png)
摘要
En 中文
We report an algorithm to identify the script of each word in a document image. We start with a bi-script scenario which is later extended to tri-script and then to eleven-script scenarios. A database of 20,000 words of different font styles and sizes has been collected and used for each script. Effectiveness of Gabor and discrete cosine transform (DCT) features has been independently evaluated using nearest neighbor, linear discriminant and support vector machines (SVM) classifiers. The combination of Gabor features with nearest neighbor or SVM classifier shows promising results; i.e., over 98% for bi-script and tri-script cases and above 89% for the eleven-script scenario. (c) 2008 Elsevier B.V. All rights reserved.
Keyword:
Gabor filter
DCT
script identification
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
IF:
3.3
论文数:
7.9K
被引数:
1.6W
机构
引用论文
Increase in the peripheral blood methylglyoxal levels in 10% of hospitalized chronic schizophrenia patients10%的住院慢性精神分裂症患者外周血甲基乙二醛水平升高

