返回
Disruptive developments in document recognition
DOI:10.1016/j.patrec.2015.11.024.png)
摘要
En 中文
Progress in optical character recognition, which underlies most applications of document processing, has been driven mainly by technological advances in microprocessors and optical sensor arrays. Software development based on algorithmic innovations appears to be reaching the point of diminishing returns. Research results, dispersed among a dozen venues, tend to lag behind commercial methodology. Some early main-line applications, like reading typescript, patents and law books, have already become obsolete. Check, postal address, and form processing are on their way out. Open source software may open up niche applications that don't generate enough revenue for commercial developers, including poorly-funded transcription of historical documents (especially genealogical records). Smartphone cameras and wearable technologies are engendering new image-based applications, but there is little evidence of widespread adoption. As document contents are integrated into a web-based continuum of data, they are likely losing even the meager individuality of discrete sheets of paper. The persistent need to create, preserve and communicate information is giving rise to entirely new genres of digital documents with a concomitant need for new approaches to document understanding. (C) 2015 Elsevier B.V. All rights reserved.
Keyword:
Optical character recognition
Document analysis
Forms processing
Semantic web
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
IF:
3.3
论文数:
7.9K
被引数:
1.6W
机构
引用论文
The Semantic Web - A new form of Web content that is meaningful to computers will unleash a revolution of new possibilities语义网-一种对计算机有意义的新形式的网络内容将引发一场新可能性的革命
SCIENTIFIC AMERICAN
IF3.1
没有更多内容

