返回
Document retrieval tolerating character recognition errors - Evaluation and application
DOI:10.1016/S0031-3203(96)00155-0.png)
摘要
En 中文
This paper presents two methods of combining character recognition with techniques for retrieving Japanese documents and also shows how these methods can be applied to textual image retrieval. Both retrieval methods are tolerant of errors that occur during the character recognition process. The basic idea is to utilize the characteristics of recognition errors. One uses a confusion matrix to generate ''equivalent'' query strings that should match erroneously recognized text. The other one searches a ''non-deterministic text'' that contains multiple candidates for ambiguous recognition results. Simulation experiments have shown that both methods can effectively combine character recognition with retrieval techniques. (C) 1997 Pattern Recognition Society. Published by Elsevier Science Ltd.
Keyword:
Japanese
character recognition
document retrieval
recognition error
confusion matrix
extended query-term method
non-deterministic text
multiple-candidate method
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

