Extraction of Virtual Baselines From Distorted Document Images Using Curvilinear Projection
Gaofeng Meng, Zuming Huang, Yonghong Song, Shiming Xiang, Chunhong Pan; Proceedings of the IEEE International Conference on Computer Vision (ICCV), 2015, pp. 3925-3933
Abstract
The baselines of a document page are a set of virtual horizontal and parallel lines, to which the printed contents of document, e.g., text lines, tables or inserted photos, are aligned. Accurate baseline extraction is of great importance in the geometric correction of curved document images. In this paper, we propose an efficient method for accurate extraction of these virtual visual cues from a curved document image. Our method comes from two basic observations that the baselines of documents do not intersect with each other and that within a narrow strip, the baselines can be well approximated by linear segments. Based upon these observations, we propose a curvilinear projection based method and model the estimation of curved baselines as a constrained sequential optimization problem. A dynamic programming algorithm is then developed to efficiently solve the problem. The proposed method can extract the complete baselines through each pixel of document images in a high accuracy. It is also scripts insensitive and highly robust to image noises, non-textual objects, image resolutions and image quality degradation like blurring and non-uniform illumination. Extensive experiments on a number of captured document images demonstrate the effectiveness of the proposed method.
Related Material
[pdf]
[
bibtex]
@InProceedings{Meng_2015_ICCV,
author = {Meng, Gaofeng and Huang, Zuming and Song, Yonghong and Xiang, Shiming and Pan, Chunhong},
title = {Extraction of Virtual Baselines From Distorted Document Images Using Curvilinear Projection},
booktitle = {Proceedings of the IEEE International Conference on Computer Vision (ICCV)},
month = {December},
year = {2015}
}