Disease named entity recognition using semisupervised learning and conditional random fields

Nichalin Suakkaphong,Zhu Zhang,Hsinchun Chen

doi:10.1002/asi.21488

Abstract

AbstractInformation extraction is an important text‐mining task that aims at extracting prespecified types of information from large text collections and making them available in structured representations such as databases. In the biomedical domain, information extraction can be applied to help biologists make the most use of their digital‐literature archives. Currently, there are large amounts of biomedical literature that contain rich information about biomedical substances. Extracting such knowledge requires a good named entity recognition technique. In this article, we combine conditional random fields (CRFs), a state‐of‐the‐art sequence‐labeling algorithm, with two semisupervised learning techniques, bootstrapping and feature sampling, to recognize disease names from biomedical literature. Two data‐processing strategies for each technique also were analyzed: one sequentially processing unlabeled data partitions and another one processing unlabeled data partitions in a round‐robin fashion. The experimental results showed the advantage of semisupervised learning techniques given limited labeled training data. Specifically, CRFs with bootstrapping implemented in sequential fashion outperformed strictly supervised CRFs for disease name recognition. The project was supported by NIH/NLM Grant R33 LM07299–01, 2002–2005.

Full Text

Paper version not known

Open DOI Link

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Disease named entity recognition using semisupervised learning and conditional random fields

Abstract

Talk to us

Similar Papers

More From: Journal of the American Society for Information Science and Technology

Lead the way for us

Journal: Journal of the American Society for Information Science and Technology	Publication Date: Feb 4, 2011
Citations: 16

Similar Papers

Biomedical Named Entity Recognition with less Supervision
Omid Ghiasvand ... Rohit J Kate
-
Omid Ghiasvand, et. al.Omid Ghiasvand ... Rohit J Kate
01 Oct 2015
01 Oct 2015

Disease named entity recognition by combining conditional random fields and bidirectional recurrent neural networks.
Qikang Wei ... Tao Chen
Database | VOL. 2016
Qikang Wei, et. al.Qikang Wei ... Tao Chen
01 Jan 2015
Database | VOL. 2016

Advanced Feature-Driven Disease Named Entity Recognition Using Conditional Random Fields
Hidayat Rahman ... Richard Segall
-
Hidayat Rahman, et. al.Hidayat Rahman ... Richard Segall
02 Oct 2016
02 Oct 2016

1 - Unified neural architecture for drug, disease, and clinical entity recognition
Sunil Kumar Sahu ... Ashish Anand
Deep Learning Techniques for Biomedical and Health Informatics | VOL. -
Sunil Kumar Sahu, et. al.Sunil Kumar Sahu ... Ashish Anand
01 Jan 2020
Deep Learning Techniques for Biomedical and Health Informatics | VOL. -

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Disease named entity recognition using semisupervised learning and conditional random fields

Abstract

Talk to us

Similar Papers

More From: Journal of the American Society for Information Science and Technology