arrow
Return

Word level multi-script identification

delete2008-07-01
delete83
PRE
AI
P
Peeta Basa Pati
A
A. G. Ramakrishnan *
DOI:10.1016/j.patrec.2008.01.027delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
We report an algorithm to identify the script of each word in a document image. We start with a bi-script scenario which is later extended to tri-script and then to eleven-script scenarios. A database of 20,000 words of different font styles and sizes has been collected and used for each script. Effectiveness of Gabor and discrete cosine transform (DCT) features has been independently evaluated using nearest neighbor, linear discriminant and support vector machines (SVM) classifiers. The combination of Gabor features with nearest neighbor or SVM classifier shows promising results; i.e., over 98% for bi-script and tri-script cases and above 89% for the eleven-script scenario. (c) 2008 Elsevier B.V. All rights reserved.
Keywords:
Gabor filter
DCT
script identification
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

Pattern Recognition Letters cover
Pattern Recognition Letters
IF:
3.3
Papers:
7.8K
Citations:
1.6W

Organization

I
indian institute of science (iisc) - bangalore
Scholars:
1.4W
Papers: 1.4W
Citations: 11