PAGE SEGMENTATION AND CLASSIFICATION

被引:111
作者
PAVLIDIS, T [1 ]
ZHOU, JY [1 ]
机构
[1] SUNY STONY BROOK,DEPT ELECT ENGN,STONY BROOK,NY 11794
来源
CVGIP-GRAPHICAL MODELS AND IMAGE PROCESSING | 1992年 / 54卷 / 06期
关键词
D O I
10.1016/1049-9652(92)90068-9
中图分类号
TP31 [计算机软件];
学科分类号
081202 ; 0835 ;
摘要
Page segmentation is the process by which a scanned page is divided into columns and blocks which are then classified as halftones, graphics, or text. Past techniques have used the fact that such parts form right rectangles for most printed material. This property is not true when the page is tilted, and the heuristics based on it fail in such cases unless a rather expensive tilt angle estimation is performed. We describe a class of techniques based on smeared run length codes that divide a page into gray and nearly white parts. Segmentation is then performed by finding connected components either by the gray elements or of the white, the latter forming white streams that partition a page into blocks of printed material. Such techniques appear quite robust in the presence of severe tilt (even greater than 10 °) and are also quite fast (about a second a page on a SPARC station for gray element aggregation). Further classification into text or halftones is based mostly on properties of the across scanlines correlation. For text correlation of adjacent scanlines tends to be quite high, but then it drops rapidly. For halftones, the correlation of adjacent scanlines is usually well below that for text, but it does not change much with distance. © 1992.
引用
收藏
页码:484 / 496
页数:13
相关论文
共 23 条
  • [1] ABELE L, 1981, 2ND P SCNAD C IM AN, P177
  • [2] Akiyama T., 1983, Transactions of the Institute of Electronics and Communication Engineers of Japan, Part D, VJ66D, P111
  • [3] BAIRD HS, 1990, 10TH P INT C PATT RE, P820
  • [4] A ROBUST ALGORITHM FOR TEXT STRING SEPARATION FROM MIXED TEXT GRAPHICS IMAGES
    FLETCHER, LA
    KASTURI, R
    [J]. IEEE TRANSACTIONS ON PATTERN ANALYSIS AND MACHINE INTELLIGENCE, 1988, 10 (06) : 910 - 918
  • [5] FLOYD R, 1975, SID 75, V75, P36
  • [6] Higashino J., 1986, Eighth International Conference on Pattern Recognition. Proceedings (Cat. No.86CH2342-4), P745
  • [7] HINDS SC, 1990, 10TH INT C PATT REC, V1, P464
  • [8] HUANG T, 1972, PICTURE BANDWIDTH CO, P231
  • [9] Kida H., 1986, Eighth International Conference on Pattern Recognition. Proceedings (Cat. No.86CH2342-4), P446
  • [10] Kubota K., 1984, Seventh International Conference on Pattern Recognition (Cat. No. 84CH2046-1), P612