Skip to main navigation Skip to search Skip to main content

Indexing and searching handwritten medical forms

Research output: Contribution to conferencePaperpeer-review

2 Scopus citations

Abstract

Extracting and reading handwritten data from medical forms is an important task in medical informatics as it paves the way for efficient archival, indexing, and retrieval. This paper addresses two important challenges: (i) extraction of handwritten text data from images of carbon copies, and (ii) intelligent use of context to reduce lexicons to make the task of handwriting recognition tractable. We have developed a smart binarization algorithm targeted to carbon copy images that outperforms methods reported in the literature. The lexicon reduction method is based on learning the medical concept, and hence the probable medical terms to be encountered in the narrative part that describes the chief complaint of the patient by training on OCR output. In our experiments, we have worked with about 600 medical forms, 20 medical concepts, and a lexicon size of 4,700. We have observed that if the concept is one of top 3 choices, the lexicon can be reduced by two-thirds on an unseen form.

Original languageEnglish
Pages73-74
Number of pages2
DOIs
StatePublished - 2006
Event7th Annual International Conference on Digital Government Research, Dg.o 2006 - San Diego, CA, United States
Duration: May 21 2006May 24 2006

Conference

Conference7th Annual International Conference on Digital Government Research, Dg.o 2006
Country/TerritoryUnited States
CitySan Diego, CA
Period05/21/0605/24/06

Keywords

  • Medical Forms Processing
  • OCR

Fingerprint

Dive into the research topics of 'Indexing and searching handwritten medical forms'. Together they form a unique fingerprint.

Cite this