e-ISSN : 0975-3397
Print ISSN : 2229-5631
Home | About Us | Contact Us

ARTICLES IN PRESS

Articles in Press

ISSUES

Current Issue
Archives

CALL FOR PAPERS

CFP 2021

TOPICS

IJCSE Topics

EDITORIAL BOARD

Editors

Indexed in

oa
 

ABSTRACT

Title : Entropy Based Texture Features Useful for Automatic Script Identification
Authors : M.C. Padma, P.A.Vijaya
Keywords : Document Processing Wavelet Packet Tree, Feature Extraction,Script Identification.
Issue Date : Mar 2010
Abstract :
In a multi script environment, a collection of documents printed in different scripts is in practice. For automatic processing of such documents through Optical Character Recognition, it is necessary to identify the script type of the document. In this paper, a novel texture-based approach is presented to identify the script type of the documents printed in three prioritized scripts - Kannada, Hindi and English, prevailed in Karnataka, an Indian state. The document images are decomposed through the Wavelet Packet Decomposition using the Haar basis function up to level two. The texture features are extracted from the sub bands of the wavelet packet decomposition. The Shannon entropy value is computed for the set of sub bands and these entropy values are combined to obtain the texture features. Experimentation conducted involved 1500 text images for learning and 1200 text images for testing. Script classification performance is analyzed using the K-nearest neighbor classifier. The average success rate is found to be 99.33%.
Page(s) : 115-120
ISSN : 0975–3397
Source : Vol. 2, Issue.2

All Rights Reserved © 2009-2024 Engg Journals Publications
Page copy protected against web site content infringement by CopyscapeCreative Commons License