A Review on Text Similarity Technique used in IR and its Application

Nitesh Pradhan,Manasi Gyanchandani,Rajesh Wadhvani

doi:10.5120/21257-4109

Abstract

With large number of documents on the web, there is a increasing need to be able to retrieve the best relevant document. There are different techniques through which we can retrieve most relevant document from the large corpus. Similarity between words, sentences, paragraphs and documents is an important component in various tasks such as information retrieval, document clustering, word-sense disambiguation, automatic essay scoring, short answer grading, machine translation and text summarization. Text similarity means user’s query text is matched with the document text and on the basis on this matching user retrieves the most relevant documents. Text similarity also plays an important role in the categorization of text as well as document. We can measure the similarity between sentences, words, paragraphs and documents to categorize them in an efficient way. On the basis of this categorization, we can retrieve the best relevant document corresponding to user’s query. This paper describes different types of similarity like lexical similarity, semantic similarity etc.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

A Review on Text Similarity Technique used in IR and its Application

Abstract

Talk to us

Similar Papers

More From: International Journal of Computer Applications

Lead the way for us

Journal: International Journal of Computer Applications	Publication Date: Jun 18, 2015
Citations: 39

Similar Papers

A Survey of Text Similarity Approaches
Wael H.Gomaa ... Aly A Fahmy
International Journal of Computer Applications | VOL. 68
Wael H.Gomaa, et. al.Wael H.Gomaa ... Aly A Fahmy
18 Apr 2013
International Journal of Computer Applications | VOL. 68

Semantic Textual Similarity Methods, Tools, and Applications: A Survey
Goutam Majumder ... Alexander Gelbukh
Computación y Sistemas | VOL. 20
Goutam Majumder, et. al.Goutam Majumder ... Alexander Gelbukh
26 Dec 2016
Computación y Sistemas | VOL. 20

Automated Identification of National Implementations of European Union Directives With Multilingual Information Retrieval Based On Semantic Textual Similarity

-

28 Mar 2019
28 Mar 2019

Semantic Textual Similarity in Japanese Clinical Domain Texts Using BERT.
Shoko Wakamiya ... Faith Wavinya Mutinda
Methods of Information in Medicine | VOL. 60
Shoko Wakamiya, et. al.Shoko Wakamiya ... Faith Wavinya Mutinda
01 Jun 2021
Methods of Information in Medicine | VOL. 60

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

A Review on Text Similarity Technique used in IR and its Application

Abstract

Talk to us

Similar Papers

More From: International Journal of Computer Applications