COMPARISON OF STEMMING AND SIMILARITY ALGORITHMS IN INDONESIAN TRANSLATED AL-QUR'AN TEXT SEARCH

Ika Oktavia Suzanti,Achmad Jauhari

doi:10.21107/kursor.v11i2.280

Abstract

The long history of information retrieval did not begin with Internet. Prior to widespread public daily use of search engines, in the 1960s information retrieval systems were discovered in commercial and intelligence applications. There are two stages in Information Retrieval in doing its main job which is to preprocessing text and to calculate similarity between term (word) and query (keyword) user searched for in a document. Stemming is final stage of pre-processing in an information retrieval system. The way stemming works is to remove affixes from a word, in form of prefixes, suffixes and insertions into form of basic word. Thus, in this paper we did compare search on information retrieval system without using stemming algorithm, using stemming Porter, Nazief & Adriani and Enhanced Confix Stripping with similarity method used is cosine similarity and dice similarity. Based on test results, text search ability on dice similarity is faster in stemming process with Porter Stemmer and ECS algorithms. While in Nazief & Adriani algorithm and without stemming, cosine similarity is faster than dice similarity.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

Journal: Jurnal Ilmiah Kursor	Publication Date: Jan 11, 2022
Citations: 1	License type: cc-by

R Discovery Prime

R Discovery Prime

COMPARISON OF STEMMING AND SIMILARITY ALGORITHMS IN INDONESIAN TRANSLATED AL-QUR'AN TEXT SEARCH

Abstract

Talk to us

Similar Papers

More From: Jurnal Ilmiah Kursor

Lead the way for us

Similar Papers

Query expansion using pseudo relevance feedback based on the bahasa version of the wikipedia dataset
Husni ... Yeni Kustiyahningsih
-
Husni, et. al. Husni ... Yeni Kustiyahningsih
01 Jan 2023
01 Jan 2023

Information Retrieval for Gujarati Language Using Cosine Similarity Based Vector Space Model
Rajnish M Rakholia ... Jatinderkumar R Saini
-
Rajnish M Rakholia, et. al.Rajnish M Rakholia ... Jatinderkumar R Saini
01 Jan 2017
01 Jan 2017

Biomedical Text Summarization: A Graph-Based Ranking Approach
Supriya Gupta ... Aakanksha Sharaff
-
Supriya Gupta, et. al.Supriya Gupta ... Aakanksha Sharaff
21 Jul 2021
21 Jul 2021

Sub-Word Indexing and Blind Relevance Feedback for English, Bengali, Hindi, and Marathi IR
Johannes Leveling ... Gareth J F Jones
ACM Transactions on Asian Language Information Processing | VOL. 9
Johannes Leveling, et. al.Johannes Leveling ... Gareth J F Jones
01 Sep 2010
ACM Transactions on Asian Language Information Processing | VOL. 9

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

COMPARISON OF STEMMING AND SIMILARITY ALGORITHMS IN INDONESIAN TRANSLATED AL-QUR'AN TEXT SEARCH

Abstract

Talk to us

Similar Papers

More From: Jurnal Ilmiah Kursor