Department of Informatics and Telecommunications, National and Kapodistrian University of Athens, Athens, Greece.
IEEE Trans Image Process. 2013 Feb;22(2):595-609. doi: 10.1109/TIP.2012.2219550. Epub 2012 Sep 18.
Document image binarization is of great importance in the document image analysis and recognition pipeline since it affects further stages of the recognition process. The evaluation of a binarization method aids in studying its algorithmic behavior, as well as verifying its effectiveness, by providing qualitative and quantitative indication of its performance. This paper addresses a pixel-based binarization evaluation methodology for historical handwritten/machine-printed document images. In the proposed evaluation scheme, the recall and precision evaluation measures are properly modified using a weighting scheme that diminishes any potential evaluation bias. Additional performance metrics of the proposed evaluation scheme consist of the percentage rates of broken and missed text, false alarms, background noise, character enlargement, and merging. Several experiments conducted in comparison with other pixel-based evaluation measures demonstrate the validity of the proposed evaluation scheme.
文档图像二值化在文档图像分析和识别管道中非常重要,因为它会影响识别过程的后续阶段。二值化方法的评估通过提供其性能的定性和定量指示,有助于研究其算法行为以及验证其有效性。本文提出了一种基于像素的历史手写/机器印刷文档图像二值化评估方法。在提出的评估方案中,使用加权方案适当修改了召回率和精度评估措施,从而减少了任何潜在的评估偏差。所提出的评估方案的其他性能指标包括断字和漏字、误报、背景噪声、字符放大和合并的百分比。与其他基于像素的评估方法进行的几次实验证明了所提出的评估方案的有效性。