Pixel-Accurate Representation and Evaluation of Page Segmentation in Document Images

Faisal Shafait, Daniel Keysers, Thomas Breuel

In: ICPR 2006, International Conference on Pattern Recognition. International Conference on Pattern Recognition (ICPR) Hong Kong IEEE Computer Society 8/2006.


This paper presents a new representation and evaluation procedure of page segmentation algorithms and analyzes six widely-used layout analysis algorithms using the procedure. The method permits a detailed analysis of the behavior of page segmentation algorithms in terms of over- and under segmentation at different layout levels, as well as determination of the geometric accuracy of the segmentation. The representation of document layouts relies on labeling each pixel according to its function in the overall segmentation, permitting pixel-accurate representation of layout information of arbitrary layouts and allowing background pixels to be classified as "don't care". Our representations can be encoded easily in standard color image formats like PNG, permitting easy interchange of segmentation results and ground truth.

icpr06PixAccRep_FsDkTmb.pdf (pdf, 174 KB )

Deutsches Forschungszentrum für Künstliche Intelligenz
German Research Center for Artificial Intelligence