Skip to main navigation Skip to search Skip to main content

Document image ground truth generation from electronic text

  • University of Maryland, College Park

Research output: Contribution to journalConference articlepeer-review

14 Scopus citations

Abstract

The problem of generating synthetic data for the training and evaluation of document analysis systems has been widely addressed in recent years. With the increased interest in processing multilingual sources, however, there is a tremendous need to be able to rapidly generate data in new languages and scripts, without the need to develop specialized systems. We have developed an approach, which uses language support of the MS Windows operating system combined with custom print drivers to render tiff images simultaneously with windows Enhanced Metafile directives. The Metafile information is parsed to generate zone, line, word, and character ground truth including location, font information and content in any language supported by Windows. The resulting images can be physically or synthetically degraded, and used for training and evaluating OCR systems. In this paper, we briefly survey related work and describe our system.

Original languageEnglish
Pages (from-to)663-666
Number of pages4
JournalProceedings - International Conference on Pattern Recognition
Volume2
StatePublished - 2004
EventProceedings of the 17th International Conference on Pattern Recognition, ICPR 2004 - Cambridge, United Kingdom
Duration: Aug 23 2004Aug 26 2004

Fingerprint

Dive into the research topics of 'Document image ground truth generation from electronic text'. Together they form a unique fingerprint.

Cite this