Abstract

Character classification in the handwritten Tamil palm-leaf manuscript is more challenging than the other document character classification due to degradation and ancient characters in the palm-leaf manuscript. In this work, RBF (Radial Basis Function) network and CART (Classification and Regression Tree) were used to classify the Tamil palm leaf segmented characters. This work consists of two phases: In the first phase, the scanned Tamil palm leaf images were preprocessed by converting them into a grayscale image and then the images were allowed to remove noise using a median filter. In the second phase, GLCM (Gray Level Co-occurrence Matrix) feature extraction method was used to extract the statistical features from the segmented characters and these features were used to train the RBF network and CART algorithm. For the RBF network, Nguyen-Widrow weight initialization technique was used to generate the weight instead of random initialization. The dataset used in this work is Kuzhanthai Pini Maruthuvam (Medicine for child-related disease). By comparing RBF using Nguyen-Widrow method with CART algorithm, RBF yields promising result of 98.4% of accuracy whereas CART produced 98.8% of accuracy for character classification. The digitization of the Tamil palm-leaf manuscript will preserve the historical secrets, traditional medicine to cure disease, healthy lifestyle, etc. It can be used in the archeological department and Tamil libraries having a palm leaf script to preserve the manuscript from degrading.

Full Text
Published version (Free)

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call