Language Identification by Using SIFT Features
Journal Title: International Journal of Advanced Research in Artificial Intelligence(IJARAI) - Year 2015, Vol 4, Issue 12
Abstract
Two novel techniques for language identification of both, machine printed and handwritten document images, are presented. Language identification is the procedure where the language of a given document image is recognized and the appropriate language label is returned. In the proposed approaches, the main body size of the characters for each document image is determined, and accordingly, a sliding window is used, in order to extract the SIFT local features. Once a large number of features have been extracted from the training set, a visual vocabulary is created, by clustering the feature space. Data clustering is performed using K-means or Gaussian Mixture Models and the Expectation - Maximization algorithm. For each document image, a Bag of Visual Words or Fisher Vector representation is constructed, using the visual vocabulary and the extracted features of the document image. Finally, a multi class Support Vector Machine classification scheme is used, to score the system. Experiments are performed on well-known databases and comparative results with another established technique, are also given.
Authors and Affiliations
Nikos Tatarakis, Ergina Kavallieratou
Using Mining Predict Relationships on the Social Media Network: Facebook (FB)
The objective of this paper is to study on the most famous social networking site Facebook and other online social media networks (OSMNs) based on the notion of relationship or friendship. This paper discussed the...
Sensor Location Problems As Test Problems Of Nonsmooth Optimization And Test Results Of A Few Nonsmooth Optimization Solvers
In this paper we address and advocate the sensor location problems and advocate them as test problems of nonsmooth optimization. These problems have easy-to-understand practical meaning and importance, easy to be e...
Geography Markup Language: GML Based Representation of Time Serie of Assimilation Data and Its Application to Animation Content Creation and Representations
Method for Geography Markup Language: GML based representation of time series of assimilation data and its application to animation content creation and representations is proposed. It is validated the proposed met...
Predicting Quality of Answer in Collaborative Q/A Community
Community Question Answering (CQA) services have emerged allowing information seekers pose their information need which is questions and receive answers from their fellow users, also participate in evaluating the q...
Static Gesture Recognition Combining Graph and Appearance Features
In this paper we propose the combination of graph-based characteristics and appearance-based descriptors such as detected edges for modeling static gestures. Initially we convolve the original image with a Gaussian...