Terminology-driven mining of biomedical literature.

Goran Nenadic,Irena Spasic,Sophia Ananiadou

doi:10.1093/bioinformatics/btg105

Goran Nenadic, Irena Spasic + Show 1 more

Open Access

https://doi.org/10.1093/bioinformatics/btg105

Copy DOI

Abstract

With an overwhelming amount of textual information in molecular biology and biomedicine, there is a need for effective literature mining techniques that can help biologists to gather and make use of the knowledge encoded in text documents. Although the knowledge is organized around sets of domain-specific terms, few literature mining systems incorporate deep and dynamic terminology processing. In this paper, we present an overview of an integrated framework for terminology-driven mining from biomedical literature. The framework integrates the following components: automatic term recognition, term variation handling, acronym acquisition, automatic discovery of term similarities and term clustering. The term variant recognition is incorporated into terminology recognition process by taking into account orthographical, morphological, syntactic, lexico-semantic and pragmatic term variations. In particular, we address acronyms as a common way of introducing term variants in biomedical papers. Term clustering is based on the automatic discovery of term similarities. We use a hybrid similarity measure, where terms are compared by using both internal and external evidence. The measure combines lexical, syntactical and contextual similarity. Experiments on terminology recognition and clustering performed on a corpus of MEDLINE abstracts recorded the precision of 98 and 71% respectively. software for the terminology management is available upon request.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Terminology-driven mining of biomedical literature.

Abstract

Talk to us

Similar Papers

More From: Bioinformatics (Oxford, England)

Lead the way for us

Journal: Bioinformatics (Oxford, England)	Publication Date: May 22, 2003
Citations: 52

Similar Papers

Terminology-driven mining of biomedical literature
Goran Nenadić ... Irena Spasić
-
Goran Nenadić, et. al.Goran Nenadić ... Irena Spasić
09 Mar 2003
09 Mar 2003

Mining term similarities from corpora
Goran Nenadic ... Sophia Ananiadou
Terminology | VOL. 10
Goran Nenadic, et. al.Goran Nenadic ... Sophia Ananiadou
10 Jun 2004
Terminology | VOL. 10

Recognizing names in biomedical texts: a machine learning approach.
Guodong Zhou ... Jian Su
Bioinformatics | VOL. 20
Guodong Zhou, et. al.Guodong Zhou ... Jian Su
05 Feb 2004
Bioinformatics | VOL. 20

Enhancing automatic term recognition through recognition of variation
Goran Nenadié ... Sophia Ananiadou
-
Goran Nenadié, et. al.Goran Nenadié ... Sophia Ananiadou
01 Jan 2004
01 Jan 2004

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Terminology-driven mining of biomedical literature.

Abstract

Talk to us

Similar Papers

More From: Bioinformatics (Oxford, England)