French-English Terminology Extraction from Comparable Corpora

Béatrice Daille,Emmanuel Morin

doi:10.1007/11562214_62

French-English Terminology Extraction from Comparable Corpora

Béatrice Daille, Emmanuel Morin

https://doi.org/10.1007/11562214_62

Copy DOI

Publication Date: Jan 1, 2005

Citations: 58

Affiliation: Laboratoire d'informatique de Nantes Atlantique

#Multi-word Terms #Single-word Terms + Show 7 more

Abstract
Full-Text PDF
Similar Papers

Abstract

This article presents a method of extracting bilingual lexica composed of single-word terms (SWTs) and multi-word terms (MWTs) from comparable corpora of a technical domain. First, this method extracts MWTs in each language, and then uses statistical methods to align single words and MWTs by exploiting the term contexts. After explaining the difficulties involved in aligning MWTs and specifying our approach, we show the adopted process for bilingual terminology extraction and the resources used in our experiments. Finally, we evaluate our approach and demonstrate its significance, particularly in relation to non-compositional MWT alignment.

Full Text