2018
Conference article  Open Access

Document bleed-through removal using sparse image inpainting

Hanif M., Tonazzini A., Savino P., Salerno E., Tsagkatakis G.

Image inpainting  Bleed-through removal  Ancient document restoration  Sparse representation 

Bleed-through is a pervasive degradation in ancient documents, caused by the ink of the opposite side of the sheet that has seeped through the paper fiber, and appears as an extra, interfering text. Bleed-through severely impairs document readability and makes it difficult to decipher the contents. Digital image restoration techniques have been successfully employed to remove or significantly reduce this distortion. The main theme is to identify the bleedthrough pixels and estimate an appropriate replacement for them, in accordance to their surrounding. This paper proposes a two-step image restoration method, exploiting information from the recto and verso images. First, based on a non-stationary linear model of the two texts overlapped in the recto-verso pair, the bleed-through pixels are identified. In the second step, a sparse representation based image inpainting technique, with a non-negative sparsity constraint, is used to find an appropriate replacement for the bleedthough pixels. Thanks to the power of dictionary learning and sparse image reconstruction methods, the natural texture of the background is well reproduced in the bleed-through areas, and even a their possible overestimation is effectively corrected, so that the original appearance of the document is preserved. The experiments are conducted on the images of a popular database of ancient documents, and the results validate the performance of the proposed method compared to the state of the art.

Source: DAS 2018 - 13th IAPR International Workshop on Document Analysis Systems, pp. 281–286, Vienna, Austria, 24-27 April 2018

Publisher: IEEE Computer Society, Los Alamitos, CA, USA


Metrics



Back to previous page
BibTeX entry
@inproceedings{oai:it.cnr:prodotti:388750,
	title = {Document bleed-through removal using sparse image inpainting},
	author = {Hanif M. and Tonazzini A. and Savino P. and Salerno E. and Tsagkatakis G.},
	publisher = {IEEE Computer Society, Los Alamitos, CA, USA},
	doi = {10.1109/das.2018.21},
	booktitle = {DAS 2018 - 13th IAPR International Workshop on Document Analysis Systems, pp. 281–286, Vienna, Austria, 24-27 April 2018},
	year = {2018}
}