Clash between Segment-level MT Error Analysis and Selected Lexical Similarity Metrics

Marija Brkic Bakaric,Lucia Nacinovic,Kristina Tonkovic

doi:10.14569/ijacsa.2020.0110506

Marija Brkic Bakaric, Lucia Nacinovic + Show 1 more

Open Access

https://doi.org/10.14569/ijacsa.2020.0110506

Copy DOI

Abstract

The aim of this paper is to evaluate the quality of popular machine translation engines on three texts of different genre in a scenario in which both source and target languages are morphologically rich. Translations are obtained from Google Translate and Microsoft Bing engines and German-Croatian is selected as the language pair. The analysis entails both human and automatic evaluation. The process of error analysis, which is time-consuming and often tiresome, is conducted in the user-friendly Windows 10 application TREAT. Prior to annotation, training is conducted in order to familiarize the annotator with MQM, which is used in the annotation task, and the interface of TREAT. The annotation guidelines elaborated with examples are provided. The evaluation is also conducted with automatic metrics BLEU and CHRF++ in order to assess their segment-level correlation with human annotations on three different levels–accuracy, mistranslation, and the total number of errors. Our findings indicate that neither the total number of errors, nor the most prominent error category and subcategory, show consistent and statistically significant segment-level correlation with the selected automatic metrics.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Clash between Segment-level MT Error Analysis and Selected Lexical Similarity Metrics

Abstract

Talk to us

Similar Papers

More From: International Journal of Advanced Computer Science and Applications

Lead the way for us

Journal: International Journal of Advanced Computer Science and Applications	Publication Date: Jan 1, 2020
License type: cc-by

Similar Papers

Human Versus Automatic Evaluation of NMT for Low-Resource Indian Language
Goutam Datta ... Nisheeth Joshi
-
Goutam Datta, et. al.Goutam Datta ... Nisheeth Joshi
01 Jan 2023
01 Jan 2023

A Survey on Evaluation Metrics for Machine Translation
Seungjun Lee ... Jungseob Lee
Mathematics | VOL. 11
Seungjun Lee, et. al.Seungjun Lee ... Jungseob Lee
16 Feb 2023
Mathematics | VOL. 11

Statistical Machine Translation Customization between Turkish and 11 Languages
Gökhan Doğru
transLogos Translation Studies Journal | VOL. 3/1
Gökhan DoğruGökhan Doğru
01 Jan 2020
transLogos Translation Studies Journal | VOL. 3/1

Machine Translation Evaluation: Manual Versus Automatic—A Comparative Study
Kaushal Kumar Maurya ... Renjith P Ravindran
-
Kaushal Kumar Maurya, et. al.Kaushal Kumar Maurya ... Renjith P Ravindran
01 Jan 2020
01 Jan 2020

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Clash between Segment-level MT Error Analysis and Selected Lexical Similarity Metrics

Abstract

Talk to us

Similar Papers

More From: International Journal of Advanced Computer Science and Applications