Sciweavers

Free Online Productivity Tools i2Speak i2Symbol i2OCR iTex2Img iWeb2Print iWeb2Shot i2Type iPdf2Split iPdf2Merge i2Bopomofo i2Arabic i2Style i2Image i2PDF iLatex2Rtf Sci2ools

158

SIGIR
2004
ACM

121views Information Technology» more SIGIR 2004»

A search engine for imaged documents in PDF files

16 years 6 days ago

A search engine for imaged documents in PDF files

Download www.comp.nus.edu.sg

Large quantities of documents in the Internet and digital libraries are simply scanned and archived in image format, many of which are packed in PDF files. The word search tool provided by Adobe Reader/Acrobat does not work for these imaged documents. In this paper, we present a search engine to deal with this issue for imaged documents in PDF files. The experimental results show an encouraging performance. Categories and Subject Descriptors H.3.3 [Information Storage and retrieval]: Information Search and Retrieval  retrieval models, search process; I.4.8 [Image Processing and Computer Vision]: Scene Analysis  Object Recognition. General Terms Algorithms, Measurement, Design, Experimentation. Keywords Imaged Document, Word searching, PDF files.

Yue Lu, Li Zhang, Chew Lim Tan

Real-time Traffic

Imaged Document | Keywords Imaged Document | PDF Files | SIGIR 2004 |

claim paper

Related Content

» Objectlevel document analysis of PDF files

» Two diet plans for fat PDF

» Xed A New Tool for eXtracting Hidden Structures from Electronic Documents

» Retrieving Metadata for Your Local Scholarly Papers

» A System for Converting PDF Documents into Structured XML Format

» XCDF A Canonical and Structured Document Format

» A Framework for the Encoding of Multilayered Documents

» SciPlore Xtract Extracting Titles from Scientific PDF Documents by Analyzing Style Informa...

» Integration of a Multilingual Keyword Extractor in a Document Management System

Post Info
More Details (n/a)

Added	30 Jun 2010
Updated	30 Jun 2010
Type	Conference
Year	2004
Where	SIGIR
Authors	Yue Lu, Li Zhang, Chew Lim Tan

Comments (0)