TY - GEN
T1 - Newspaper image understanding
AU - Govindaraju, Venu
AU - Lam, Stephen W.
AU - Niyogi, Debashish
AU - Sher, David B.
AU - Srihari, Rohini
AU - Srihari, Sargur N.
AU - Wang, Dacheng
N1 - Publisher Copyright:
© Springer-Verlag Berlin Heidelberg 1990.
PY - 1990
Y1 - 1990
N2 - Understanding printed documents such as newspapers is a common intelligent activity of humans. Making a computer perform the task of analyzing a newspaper image and derive useful high-level representations requires the development and integration of techniques in several areas, including pattern recognition, computer vision, language understanding and artificial intelligence. We describe the organization and several components of a newspaper image undertanding system that begins with digitized images of newspaper pages and produces symbolic representations at several different levels. Such representations include: the visual sketch (connected components extracted from the background), physical layout (spatial extents of blocks corresponding to text, half-tones, graphics), logical layout (organization of story components), block primitives (e.g., recognized characters and words in text blocks, lines in graphics, faces in photographs, etc.), and semantic nets corresponding to photographic and textual blocks (individually, as well as grouped together as stories). We describe algorithms for deriving several of the representations and describe the interaction of different modules.
AB - Understanding printed documents such as newspapers is a common intelligent activity of humans. Making a computer perform the task of analyzing a newspaper image and derive useful high-level representations requires the development and integration of techniques in several areas, including pattern recognition, computer vision, language understanding and artificial intelligence. We describe the organization and several components of a newspaper image undertanding system that begins with digitized images of newspaper pages and produces symbolic representations at several different levels. Such representations include: the visual sketch (connected components extracted from the background), physical layout (spatial extents of blocks corresponding to text, half-tones, graphics), logical layout (organization of story components), block primitives (e.g., recognized characters and words in text blocks, lines in graphics, faces in photographs, etc.), and semantic nets corresponding to photographic and textual blocks (individually, as well as grouped together as stories). We describe algorithms for deriving several of the representations and describe the interaction of different modules.
UR - https://www.scopus.com/pages/publications/3342951397
U2 - 10.1007/BFb0018395
DO - 10.1007/BFb0018395
M3 - Conference contribution
AN - SCOPUS:3342951397
SN - 9783540528500
T3 - Lecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics)
SP - 375
EP - 384
BT - Knowledge Based Computer Systems - International Conference KBCS 1989, Proceedings
A2 - Ramani, S.
A2 - Chandrasekar, R.
A2 - Anjaneyulu, K.S.R.
PB - Springer Verlag
T2 - 2nd International Conference on Knowledge Based Computer Systems, KBCS 1989
Y2 - 11 December 1989 through 13 December 1989
ER -