A Review on Text Similarity Technique used in IR and its Application

  • Pradhan N
  • Gyanchandani M
  • Wadhvani R
N/ACitations
Citations of this article
90Readers
Mendeley users who have this article in their library.

Abstract

With large number of documents on the web, there is a increasing need to be able to retrieve the best relevant document. There are different techniques through which we can retrieve most relevant document from the large corpus. Similarity between words, sentences, paragraphs and documents is an important component in various tasks such as information retrieval, document clustering, word-sense disambiguation, automatic essay scoring, short answer grading, machine translation and text summarization. Text similarity means user's query text is matched with the document text and on the basis on this matching user retrieves the most relevant documents. Text similarity also plays an important role in the categorization of text as well as document. We can measure the similarity between sentences, words, paragraphs and documents to categorize them in an efficient way. On the basis of this categorization, we can retrieve the best relevant document corresponding to user's query. This paper describes different types of similarity like lexical similarity, semantic similarity etc. General Term Text Similarity, Text Mining, Text Summarization Keyword Text similarity, Lexical similarity, semantic similarity, Corpus based similarity and Knowledge based similarity.

Cite

CITATION STYLE

APA

Pradhan, N., Gyanchandani, M., & Wadhvani, R. (2015). A Review on Text Similarity Technique used in IR and its Application. International Journal of Computer Applications, 120(9), 29–34. https://doi.org/10.5120/21257-4109

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free