Indonesian part of speech tagging using maximum entropy markov model on Indonesian manually tagged corpus

Denis Eka Cahyani,Winda Mustikaningtyas

doi:10.11591/ijai.v11.i1.pp336-344

Abstract

This research discusses the development of a part of speech (POS) tagging system to solve the problem of word ambiguity. This paper presents a new method, namely maximum entropy markov model (MEMM) to solve word ambiguity on the Indonesian dataset. A manually labeled “Indonesian manually tagged corpus” was used as data. Furthermore, the corpus is processed using the entropy formula to obtain the weight of the value of the word being searched for, then calculating it into the MEMM Bigram and MEMM Trigram algorithms with the previously obtained rules to determine the part of speech (POS) tag that has the highest probability. The results obtained show POS tagging using the MEMM method has advantages over the methods used previously which used the same data. This paper improves a performance evaluation of research previously. The resulting average accuracy is 83.04% for the MEMM Bigram algorithm and 86.66% for the MEMM Trigram. The MEMM Trigram algorithm is better than the MEMM Bigram algorithm.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

Journal: IAES International Journal of Artificial Intelligence (IJ-AI)	Publication Date: Mar 1, 2022
Citations: 1	License type: CC BY-SA 4.0

R Discovery Prime

R Discovery Prime

Indonesian part of speech tagging using maximum entropy markov model on Indonesian manually tagged corpus

Abstract

Talk to us

Similar Papers

More From: IAES International Journal of Artificial Intelligence (IJ-AI)

Lead the way for us

Similar Papers

Training MEMM with PSO: A Tool for Part-of-Speech Tagging
Lei La ... Qiao Guo
Journal of Software | VOL. 7
Lei La, et. al.Lei La ... Qiao Guo
11 Jan 2012
Journal of Software | VOL. 7

CRF Models for Tamil Part of Speech Tagging and Chunking
S Lakshmana Pandian ... T V Geetha
-
S Lakshmana Pandian, et. al.S Lakshmana Pandian ... T V Geetha
01 Jan 2009
01 Jan 2009

Part-of-Speech Tagging of Odia Language Using Statistical and Deep Learning Based Approaches
Tusarkanta Dalai ... Pankaj K Sa
ACM Transactions on Asian and Low-Resource Language Information Processing | VOL. 22
Tusarkanta Dalai, et. al.Tusarkanta Dalai ... Pankaj K Sa
16 Jun 2023
ACM Transactions on Asian and Low-Resource Language Information Processing | VOL. 22

The study of a nonstationary maximum entropy Markov model and its application on the pos-tagging task
Jinghui Xiao ... Bingquan Liu
ACM Transactions on Asian Language Information Processing | VOL. 6
Jinghui Xiao, et. al.Jinghui Xiao ... Bingquan Liu
01 Sep 2007
ACM Transactions on Asian Language Information Processing | VOL. 6

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Indonesian part of speech tagging using maximum entropy markov model on Indonesian manually tagged corpus

Abstract

Talk to us

Similar Papers

More From: IAES International Journal of Artificial Intelligence (IJ-AI)