Improved Named Entity Recognition using Machine Translation-based Cross-lingual Information

Sandipan Dandapat,Andy Way

doi:10.13053/cys-20-3-2468

Improved Named Entity Recognition using Machine Translation-based Cross-lingual Information

Sandipan Dandapat, Andy Way

Open Access

https://doi.org/10.13053/cys-20-3-2468

Copy DOI

Journal: Computación y Sistemas	Publication Date: Sep 30, 2016
Citations: 9	License type: other-oa

#Cross-lingual Features #On-line Machine Translation System + Show 8 more

Abstract
Full-Text PDF
Similar Papers

Abstract

In this paper, we describe a technique to improve named entity recognition in a resource-poor language (Hindi) by using cross-lingual information. We use an on-line machine translation system and a separate word alignment phase to find the projection of each Hindi word into the translated English sentence. We estimate the cross-lingual features using an English named entity recognizer and the alignment information. We use these cross-lingual features in a support vector machine-based classifier. The use of cross-lingual features improves F 1 score by 2.1 points absolute (2.9% relative) over a good-performing baseline model.

Full Text

Paper version not known

Open DOI Link

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

Similar Papers

Paper Title

Journal

Date

Author

View more papers

More From: Computación y Sistemas

Paper Title

Journal

Date

Author

View more papers

Disclaimer: All third-party content on this website/platform is and will remain the property of their respective owners and is provided on "as is" basis without any warranties, express or implied. Use of third-party content does not indicate any affiliation, sponsorship with or endorsement by them. Any references to third-party content is to identify the corresponding services and shall be considered fair use under The CopyrightLaw.