Handwritten Numeral Databases of Indian Scripts and Multistage Recognition of Mixed Numerals
- 18 April 2008
- journal article
- Published by Institute of Electrical and Electronics Engineers (IEEE) in IEEE Transactions on Pattern Analysis and Machine Intelligence
- Vol. 31 (3), 444-457
- https://doi.org/10.1109/tpami.2008.88
Abstract
This article primarily concerns the problem of isolated handwritten numeral recognition of major Indian scripts. The principal contributions presented here are (a) pioneering development of two databases for handwritten numerals of two most popular Indian scripts, (b) a multistage cascaded recognition scheme using wavelet based multiresolution representations and multilayer perceptron classifiers and (c) application of (b) for the recognition of mixed handwritten numerals of three Indian scripts Devanagari, Bangla and English. The present databases include respectively 22,556 and 23,392 handwritten isolated numeral samples of Devanagari and Bangla collected from real-life situations and these can be made available free of cost to researchers of other academic Institutions. In the proposed scheme, a numeral is subjected to three multilayer perceptron classifiers corresponding to three coarse-to-fine resolution levels in a cascaded manner. If rejection occurred even at the highest resolution, another multilayer perceptron is used as the final attempt to recognize the input numeral by combining the outputs of three classifiers of the previous stages. This scheme has been extended to the situation when the script of a document is not known a priori or the numerals written on a document belong to different scripts. Handwritten numerals in mixed scripts are frequently found in Indian postal mails and table-form documents.This publication has 47 references indexed in Scilit:
- Fuzzy model based recognition of handwritten numeralsPattern Recognition, 2007
- Handwritten digit recognition: benchmarking of state-of-the-art techniquesPattern Recognition, 2003
- Handwritten numeral recognition using gradient and curvature of gray scale imagePattern Recognition, 2002
- A complete printed Bangla OCR systemPattern Recognition, 1998
- Feature extraction methods for character recognition-A surveyPattern Recognition, 1996
- Recognition of handwritten numerals with multiple feature and multistage classifierPattern Recognition, 1995
- Nonlinear shape normalization methods for the recognition of large-set handwritten charactersPattern Recognition, 1994
- A nonlinear normalization method for handprinted kanji character recognition—line density equalizationPattern Recognition, 1990
- Rule based contextual post-processing for devanagari text recognitionPattern Recognition, 1987
- Machine recognition of constrained hand printed devanagariPattern Recognition, 1977