Automatic recognition of multi-word terms:. the C-value/NC-value method

Katerina Frantzi,Hideki Mima,Sophia Ananiadou

doi:10.1007/s007999900023

Automatic recognition of multi-word terms:. the C-value/NC-value method

Katerina Frantzi, Hideki Mima + Show 1 more

Open Access

https://doi.org/10.1007/s007999900023

Copy DOI

Journal: International Journal on Digital Libraries	Publication Date: Aug 1, 2000
Citations: 742

Affiliation: University of Manchester, The University of Tokyo, Tokyo University of Science

#Extraction Of Terms #Automatic Extraction Of Terms + Show 8 more

Abstract
Full-Text PDF
Similar Papers

Abstract

Technical terms (henceforth called terms ), are important elements for digital libraries. In this paper we present a domain-independent method for the automatic extraction of multi-word terms, from machine-readable special language corpora. The method, (C-value/NC-value ), combines linguistic and statistical information. The first part, C-value, enhances the common statistical measure of frequency of occurrence for term extraction, making it sensitive to a particular type of multi-word terms, the nested terms. The second part, NC-value, gives: 1) a method for the extraction of term context words (words that tend to appear with terms); 2) the incorporation of information from term context words to the extraction of terms.

Full Text

Paper version not known

Open DOI Link

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

Similar Papers

Paper Title

Journal

Date

Author

View more papers

More From: International Journal on Digital Libraries

Paper Title

Journal

Date

Author

View more papers

Disclaimer: All third-party content on this website/platform is and will remain the property of their respective owners and is provided on "as is" basis without any warranties, express or implied. Use of third-party content does not indicate any affiliation, sponsorship with or endorsement by them. Any references to third-party content is to identify the corresponding services and shall be considered fair use under The CopyrightLaw.