A New Binarization Algorithm for Historical Documents
AbstractMonochromatic documents claim for much less computer bandwidth for network transmission and storage space than their color or even grayscale equivalent. The binarization of historical documents is far more complex than recent ones as paper aging, color, texture, translucidity, stains, back-to-front interference, kind and color of ink used in handwriting, printing process, digitalization process, etc. are some of the factors that affect binarization. This article presents a new binarization algorithm for historical documents. The new global filter proposed is performed in four steps: filtering the image using a bilateral filter, splitting image into the RGB components, decision-making for each RGB channel based on an adaptive binarization method inspired by Otsu’s method with a choice of the threshold level, and classification of the binarized images to decide which of the RGB components best preserved the document information in the foreground. The quantitative and qualitative assessment made with 23 binarization algorithms in three sets of “real world” documents showed very good results. View Full-Text
Scifeed alert for new publicationsNever miss any articles matching your research from any publisher
- Get alerts for new papers matching your research
- Find out the new papers from selected authors
- Updated daily for 49'000+ journals and 6000+ publishers
- Define your Scifeed now
Almeida, M.; Lins, R.D.; Bernardino, R.; Jesus, D.; Lima, B. A New Binarization Algorithm for Historical Documents. J. Imaging 2018, 4, 27.
Almeida M, Lins RD, Bernardino R, Jesus D, Lima B. A New Binarization Algorithm for Historical Documents. Journal of Imaging. 2018; 4(2):27.Chicago/Turabian Style
Almeida, Marcos; Lins, Rafael D.; Bernardino, Rodrigo; Jesus, Darlisson; Lima, Bruno. 2018. "A New Binarization Algorithm for Historical Documents." J. Imaging 4, no. 2: 27.
Note that from the first issue of 2016, MDPI journals use article numbers instead of page numbers. See further details here.