Pattern Recognition, International Conference on (2006)
Aug. 20, 2006 to Aug. 24, 2006
DOI Bookmark: http://doi.ieeecomputersociety.org/10.1109/ICPR.2006.934
Faisal Shafait , Technical University of Kaiserslautern, Germany
Daniel Keysers , Technical University of Kaiserslautern, Germany
Thomas M. Breuel , Technical University of Kaiserslautern, Germany
This paper presents a new representation and evaluation procedure of page segmentation algorithms and analyzes six widely-used layout analysis algorithms using the procedure. The method permits a detailed analysis of the behavior of page segmentation algorithms in terms of over- and undersegmentation at different layout levels, as well as determination of the geometric accuracy of the segmentation. The representation of document layouts relies on labeling each pixel according to its function in the overall segmentation, permitting pixel-accurate representation of layout information of arbitrary layouts and allowing background pixels to be classified as "don?t care". Our representations can be encoded easily in standard color image formats like PNG, permitting easy interchange of segmentation results and ground truth.
F. Shafait, D. Keysers and T. M. Breuel, "Pixel-Accurate Representation and Evaluation of Page Segmentation in Document Images," 2006 18th International Conference on Pattern Recognition(ICPR), Hong Kong, 2006, pp. 872-875.