A semi-supervised approach to extract pharmacogenomics-specific drug–gene pairs from biomedical literature for personalized medicine

Rong Xu,Quanqiu Wang

doi:10.1016/j.jbi.2013.04.001

Rong Xu, Quanqiu Wang

Open Access

https://doi.org/10.1016/j.jbi.2013.04.001

Copy DOI

Journal: Journal of Biomedical Informatics	Publication Date: Apr 6, 2013
Citations: 31	License type: publisher-specific-oa

Affiliation: Case Western Reserve University

Abstract

Personalized medicine is to deliver the right drug to the right patient in the right dose. Pharmacogenomics (PGx) is to identify genetic variants that may affect drug efficacy and toxicity. The availability of a comprehensive and accurate PGx-specific drug–gene relationship knowledge base is important for personalized medicine. However, building a large-scale PGx-specific drug–gene knowledge base is a difficult task. In this study, we developed a bootstrapping, semi-supervised learning approach to iteratively extract and rank drug–gene pairs according to their relevance to drug pharmacogenomics. Starting with a single PGx-specific seed pair and 20 million MEDLINE abstracts, the extraction algorithm achieved a precision of 0.219, recall of 0.368 and F1 of 0.274 after two iterations, a significant improvement over the results of using non-PGx-specific seeds (precision: 0.011, recall: 0.018, and F1: 0.014) or co-occurrence (precision: 0.015, recall: 1.000, and F1: 0.030). After the extraction step, the ranking algorithm further improved the precision from 0.219 to 0.561 for top ranked pairs. By comparing to a dictionary-based approach with PGx-specific gene lexicon as input, we showed that the bootstrapping approach has better performance in terms of both precision and F1 (precision: 0.251 vs. 0.152, recall: 0.396 vs. 0.856 and F1: 0.292 vs. 0.254). By integrative analysis using a large drug adverse event database, we have shown that the extracted drug–gene pairs strongly correlate with drug adverse events. In conclusion, we developed a novel semi-supervised bootstrapping approach for effective PGx-specific drug–gene pair extraction from large number of MEDLINE articles with minimal human input.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

A semi-supervised approach to extract pharmacogenomics-specific drug–gene pairs from biomedical literature for personalized medicine

Abstract

Talk to us

Similar Papers

More From: Journal of Biomedical Informatics

Lead the way for us

Similar Papers

Pharmacogenomics of Drug-Metabolizing Enzymes and Drug Transporters in Chemotherapy
Tessa M Bosch
-
Tessa M BoschTessa M Bosch
01 Jan 2008
01 Jan 2008

Ambulatory Computerized Prescribing and Preventable Adverse Drug Events
Joseph Marcus Overhage ... Michael D Murray
Journal of Patient Safety | VOL. 12
Joseph Marcus Overhage, et. al.Joseph Marcus Overhage ... Michael D Murray
01 Jun 2016
Journal of Patient Safety | VOL. 12

Extraction of Information Related to Drug Safety Surveillance From Electronic Health Record Notes: Joint Modeling of Entities and Relations Using Knowledge-Aware Neural Attentive Models.
Bharath Dandala ... Jennifer J Liang
JMIR Medical Informatics | VOL. 8
Bharath Dandala, et. al.Bharath Dandala ... Jennifer J Liang
10 Jul 2020
JMIR Medical Informatics | VOL. 8

Preventability of adverse drug events involving multiple drugs using publicly available clinical decision support tools
Adam Wright ... Joshua Feblowitz
American Journal of Health-System Pharmacy | VOL. 69
Adam Wright, et. al.Adam Wright ... Joshua Feblowitz
01 Feb 2012
American Journal of Health-System Pharmacy | VOL. 69

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

A semi-supervised approach to extract pharmacogenomics-specific drug–gene pairs from biomedical literature for personalized medicine

Abstract

Talk to us

Similar Papers

More From: Journal of Biomedical Informatics