Efficient Hybrid Semantic Text Similarity using Wordnet and a Corpus

Abstract

Text similarity plays an important role in natural language processing tasks such as answering questions and summarizing text. At present, state-of-the-art text similarity algorithms rely on inefficient word pairings and/or knowledge derived from large corpora such as Wikipedia. This article evaluates previous word similarity measures on benchmark datasets and then uses a hybrid word similarity in a novel text similarity measure (TSM). The proposed TSM is based on information content and WordNet semantic relations. TSM includes exact word match, the length of both sentences in a pair, and the maximum similarity between one word and the compared text. Compared with other well-known measures, results of TSM are surpassing or comparable with the best algorithms in the literature.

Authors and Affiliations

Issa Atoum, Ahmed Otoom

Keywords

Related Articles

Bidirectional WDM-Radio over Fiber System with Sub Carrier Multiplexing Using a Reflective SOA and Cyclic AWGs 

A bidirectional SCM-WDM RoF network using a reflective semiconductor optical amplifier (RSOA) and cyclic arrayed waveguide gratings (AWGs) was proposed and demonstrated. The purposed RoF network utilizes Sub Carrier Mult...

A Survey of Cloud Migration Methods: A Comparison and Proposition

Along with the significant advantages of cloud computing paradigm, the number of enterprises, which expect to move a legacy system towards a cloud, is steadily increasing. Unfortunately, this move is not straightforward....

Development of Home Network Sustainable Interface Tools

The home network has become a norm in today's life. Previous studies have shown that home network management is a problem for users who are not in the field of network technology. The existing network management tools ar...

Comparison of Task Scheduling Algorithms in Cloud Environment

The enhanced form of client-server, cluster and grid computing is termed as Cloud Computing. The cloud users can virtually access the resources over the internet. Task submitted by cloud users are responsible for efficie...

Performance Analysis of Machine Learning Algorithms for Missing Value Imputation

Data mining requires a pre-processing task in which the data are prepared, cleaned, integrated, transformed, reduced and discretized for ensuring the quality. Missing values is a universal problem in many research domain...

Download PDF file
  • EP ID EP149411
  • DOI 10.14569/IJACSA.2016.070917
  • Views 113
  • Downloads 0

How To Cite

Issa Atoum, Ahmed Otoom (2016). Efficient Hybrid Semantic Text Similarity using Wordnet and a Corpus. International Journal of Advanced Computer Science & Applications, 7(9), 124-130. https://europub.co.uk/articles/-A-149411