Part-of-Speech Tags Guide Low-Resource Machine Translation

Zaokere Kadeer,Nian Yi,Aishan Wumaier

doi:10.3390/electronics12163401

Abstract

Neural machine translation models are guided by loss function to select source sentence features and generate results close to human annotation. When the data resources are abundant, neural machine translation models can focus on the features used to produce high-quality translations. These features include POS or other grammatical features. However, models cannot focus precisely on these features when data resources are limited. The reason is that the lack of samples makes the model overfit before considering these features. Previous works have enriched the features by integrating source POS or multitask methods. However, these methods only utilize the source POS or produce translations by introducing the generated target POS. We propose introducing POS information based on multitask methods and reconstructors. We obtain the POS tags by the additional encoder and decoder and compute the corresponding loss function. These loss functions are used with the loss function of machine translation to optimize the parameters of the entire model, which makes the model pay attention to POS features. The POS features focused on by models will guide the translation process and alleviate the problem that models cannot focus on the POS features in the case of low resources. Experiments on multiple translation tasks show that the method improves 0.4∼1 BLEU compared with the baseline model on different translation tasks.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Part-of-Speech Tags Guide Low-Resource Machine Translation

Abstract

Talk to us

Similar Papers

More From: Electronics

Lead the way for us

Journal: Electronics	Publication Date: Aug 10, 2023
License type: CC BY 4.0

Similar Papers

A Multitask-Based Neural Machine Translation Model with Part-of-Speech Tags Integration for Arabic Dialects
Laith H Baniata ... Seong-Bae Park
Applied Sciences | VOL. 8
Laith H Baniata, et. al.Laith H Baniata ... Seong-Bae Park
05 Dec 2018
Applied Sciences | VOL. 8

What do Neural Machine Translation Models Learn about Morphology?
Yonatan Belinkov ... Nadir Durrani
-
Yonatan Belinkov, et. al.Yonatan Belinkov ... Nadir Durrani
01 Jan 2017
01 Jan 2017

A Document-Level Neural Machine Translation Model with Dynamic Caching Guided by Theme-Rheme Information
Yiqi Tong ... Xiaodong Shi
-
Yiqi Tong, et. al.Yiqi Tong ... Xiaodong Shi
01 Jan 2020
01 Jan 2020

Adversarial Subword Regularization for Robust Neural Machine Translation
Jungsoo Park ... Jaewoo Kang
-
Jungsoo Park, et. al.Jungsoo Park ... Jaewoo Kang
01 Jan 2020
01 Jan 2020

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Part-of-Speech Tags Guide Low-Resource Machine Translation

Abstract

Talk to us

Similar Papers

More From: Electronics