A copy of this work was available on the public web and has been preserved in the Wayback Machine. The capture dates from 2011; you can also visit the original URL.
The file type is application/pdf
.
An automatic linking service of document images reducing the effects of OCR errors with latent semantics
2010
Proceedings of the 2010 ACM Symposium on Applied Computing - SAC '10
Robust Information Retrieval (IR) systems have been demanded due to the widespread and multipurpose use of document images, and the high number of document images repositories available nowadays. This paper presents a novel approach to support the automatic generation of relationships among document images by exploiting Latent Semantic Indexing (LSI) and Optical Character Recognition (OCR). The LinkDI service extracts and indexes document images content, obtains its latent semantics, and
doi:10.1145/1774088.1774092
dblp:conf/sac/NetoGBPM10
fatcat:xqujwg56qvhgrm7zev6467qfna