Development of a classification scheme for disease-related enzyme information

Carola Söhngen,Dietmar Schomburg,Antje Chang

doi:10.1186/1471-2105-12-329

Carola Söhngen, Dietmar Schomburg + Show 1 more

Open Access

https://doi.org/10.1186/1471-2105-12-329

Copy DOI

Journal: BMC Bioinformatics	Publication Date: Aug 9, 2011
Citations: 40	License type: cc-by

Affiliation: Technische Universität Braunschweig

Abstract

BackgroundBRENDA (BRaunschweig ENzyme DAtabase, http://www.brenda-enzymes.org) is a major resource for enzyme related information. First and foremost, it provides data which are manually curated from the primary literature. DRENDA (Disease RElated ENzyme information DAtabase) complements BRENDA with a focus on the automatic search and categorization of enzyme and disease related information from title and abstracts of primary publications. In a two-step procedure DRENDA makes use of text mining and machine learning methods.ResultsCurrently enzyme and disease related references are biannually updated as part of the standard BRENDA update. 910,897 relations of EC-numbers and diseases were extracted from titles or abstracts and are included in the second release in 2010. The enzyme and disease entity recognition has been successfully enhanced by a further relation classification via machine learning. The classification step has been evaluated by a 5-fold cross validation and achieves an F1 score between 0.802 ± 0.032 and 0.738 ± 0.033 depending on the categories and pre-processing procedures. In the eventual DRENDA content every category reaches a classification specificity of at least 96.7% and a precision that ranges from 86-98% in the highest confidence level, and 64-83% for the smallest confidence level associated with higher recall.ConclusionsThe DRENDA processing chain analyses PubMed, locates references with disease-related information on enzymes and categorises their focus according to the categories causal interaction, therapeutic application, diagnostic usage and ongoing research. The categorisation gives an impression on the focus of the located references. Thus, the relation categorisation can facilitate orientation within the rapidly growing number of references with impact on diseases and enzymes. The DRENDA information is available as additional information in BRENDA.

Highlights

BRENDA (BRaunschweig ENzyme DAtabase, http://www.brenda-enzymes.org) is a major resource for enzyme related information
Enzyme and Disease related references The content of DRENDA focuses on information on enzymes and ignores other proteins which may be of biomedical relevance in a physiological and pathophysiological context. This selection is caused by the available extended dictionaries of enzyme names and on the close relationship of DRENDA as a text mining derived knowledge source which contributes to the expert annotated information of BRENDA [7]
DRENDA provides broad information on enzymes and diseases arranged according to Enzyme class (EC) number and disease name

Summary

Introduction

BRENDA (BRaunschweig ENzyme DAtabase, http://www.brenda-enzymes.org) is a major resource for enzyme related information. First and foremost, it provides data which are manually curated from the primary literature. On the one hand the work of manual expert information extraction is time-consuming and expensive and on the other hand computational power gets cheaper and the growth of publications in the biomedical field is tremendous. Applications in this area quickly gain more importance. The data are extracted from literature abstracts using text-mining procedures

Methods

Results

Conclusion

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Development of a classification scheme for disease-related enzyme information

Abstract

Highlights

Summary

Talk to us

Similar Papers

More From: BMC Bioinformatics

Lead the way for us

Similar Papers

A disease-specific language representation model for cerebrovascular disease research
Ching-Heng Lin ... Yang C Fann
Computer Methods and Programs in Biomedicine | VOL. 211
Ching-Heng Lin, et. al.Ching-Heng Lin ... Yang C Fann
30 Sep 2021
Computer Methods and Programs in Biomedicine | VOL. 211

A cross-sectional study on current status of rare disease related health information based on WeChat official accounts in China
L Xu ... X F Lai
Zhonghua liu xing bing xue za zhi = Zhonghua liuxingbingxue zazhi | VOL. 41
L Xu, et. al.L Xu ... X F Lai
10 Mar 2020
Zhonghua liu xing bing xue za zhi = Zhonghua liuxingbingxue zazhi | VOL. 41

Knowledge, attitudes and quality of life of type 2 diabetes patients in Saudi Arabia
Ibrahim Suliman Alaboudi ... Asim Hassan
Saudi Pharmaceutical Journal | VOL. -
Ibrahim Suliman Alaboudi, et. al.Ibrahim Suliman Alaboudi ... Asim Hassan
01 Aug 2014
Saudi Pharmaceutical Journal | VOL. -

A Novel Drug Repositioning Approach Based on Collaborative Metric Learning.
Huimin Luo ... Fang-Xiang Wu
IEEE/ACM Transactions on Computational Biology and Bioinformatics | VOL. 18
Huimin Luo, et. al.Huimin Luo ... Fang-Xiang Wu
02 Jul 2019
IEEE/ACM Transactions on Computational Biology and Bioinformatics | VOL. 18

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Development of a classification scheme for disease-related enzyme information

Abstract

Highlights

Summary

Talk to us

Similar Papers

More From: BMC Bioinformatics