Line segmentation challenges in tamil language palm leaf manuscripts

0Citations
Citations of this article
3Readers
Mendeley users who have this article in their library.
Get full text

Abstract

The process of an Optical Character Recognition (OCR) for ancient hand written documents or palm leaf manuscripts is done by means of four phases. The four phases are ‘line segmentation’, ‘word segmentation’, ‘character segmentation’, and ‘character recognition’. The colour image of palm leaf manuscripts are changed into binary images by using various pre-processing methods. The first phase of an OCR might break through the hurdles of touching lines and overlapping lines. The character recognition becomes futile when the line segmentation is erroneous. In Tamil language palm leaf manuscript recognition, there are only a handful of line segmentation methods. Moreover, the available methods are not viable to meet the required standards. This article is proposed to fill the lacuna in terms of the methods necessary for line segmentation in Tamil language document analysis. The method proposed compares its efficiency with the line segmentation algorithms work on binary images such as the Adaptive Partial Projection (APP) and A* Path Planning (A*PP). The tools and criteria of evaluation metrics are measured from ICDAR 2013 Handwriting Segmentation Contest.

Cite

CITATION STYLE

APA

Spurgen Ratheash, R., & Mohamed Sathik, M. (2019). Line segmentation challenges in tamil language palm leaf manuscripts. International Journal of Innovative Technology and Exploring Engineering, 9(1), 2363–2367. https://doi.org/10.35940/ijitee.L3159.119119

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free