Entropy Based Texture Features Useful for Automatic Script Identification
Journal Title: International Journal on Computer Science and Engineering - Year 2010, Vol 2, Issue 2
Abstract
In a multi script environment, a collection of documents printed in different scripts is in practice. For automatic processing of such documents through Optical Character Recognition, it is necessary to identify the script type of the document. In this paper, a novel texture-based approach is presented to identify the script type of the documents printed in three prioritized scripts - Kannada, Hindi and English, prevailed in Karnataka, an Indian state. The document images are decomposed through the Wavelet Packet Decomposition using the Haar basis function up to level two. The texture features are extracted from the sub bands of the wavelet packet decomposition. The Shannon entropy value is computed for the set of sub bands and these entropy values are combined to obtain the texture features. Experimentation conducted involved 1500 text images for learning and 1200 text images for testing. Script classification performance is analyzed using the Knearest neighbor classifier. The average success rate is found to be 99.33%.
Authors and Affiliations
M. C. Padma , P. A. Vijaya
Web Service Security Through Business Logic
Service Computing has recently gained significant momentum to enable the IT services and computing technology to perform business services more efficiently and effectively. Presently the service computing concentrates on...
ACO Based Feature Subset Selection for Multiple k-Nearest Neighbor Classifiers
The k-nearest neighbor (k-NN) is one of the most popular algorithms used for classification in various fields of pattern recognition & data mining problems. In k-nearest neighbor classification, the result of a new i...
Performance of SIFT based Video Retrieval
Video has become an important element of multimedia computing and communication environments, with applications as varied as broadcasting, education, publishing and military intelligence. In Video Retrieval system, each...
Study of Data Mining Approach for Mobile Computing Environment
Efficient Data mining Techniques are required to discover useful Information andknowledge. This is due to the effective involvement of computers and the improvement inDatabase Technology which has provided large Data. Th...
An Approach to Active Queue Management in Computer Network
Active queue management is a key technique for reducing the packet drop rate in the internet. This packet dropping mechanism is used in a router to minimize congestion when the packets are dropped before queue gets full...