Abstract
In this paper, we describe a compression and representation scheme which exploits the component-level redundancy found within a document image. The approach identifies patterns which appear repeatedly, represents similar patterns with a single prototype, stores the location of pattern instances, and codes the residuals between the prototypes and the pattern instances. Using a novel encoding scheme, we provide a representation that facilitates scalable lossy compression and progressive transmission and supports document image analysis in the compressed domain. We motivate the approach, provide details of the encoding procedures, report compression results, and describe a class of document image understanding tasks that operate on the compressed representation.
| Original language | English |
|---|---|
| Pages (from-to) | 335-349 |
| Number of pages | 15 |
| Journal | Computer Vision and Image Understanding |
| Volume | 70 |
| Issue number | 3 |
| DOIs | |
| State | Published - Jun 1998 |
Fingerprint
Dive into the research topics of 'Symbolic Compression and Processing of Document Images'. Together they form a unique fingerprint.Cite this
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver