Estimating the prevalence of diabetic retinopathy in electronic health records with massive missing labels

Ye Liang,Ru Wang,Yuchen Wang,Tieming Liu

doi:10.1016/j.ibmed.2024.100154

Abstract

ObjectiveThe paper aims to address the problem of massive unlabeled patients in electronic health records (EHR) who potentially have undiagnosed diabetic retinopathy (DR). It is desired to estimate the actual DR prevalence in EHR with 96 % missing labels. Materials and methodsThe Cerner Health Facts data are used in the study, with 3749 labeled DR patients and 97,876 unlabeled diabetic patients. This extensive dataset spans the demographics of the United States over the past two decades. We implemented state-of-art positive-unlabeled learning methods, including ensemble-based support vector machine, ensemble-based random forest, and Bayesian finite mixture modeling. ResultsThe estimated DR prevalence in the population represented by Cerner EHR is approximately 25 % and the classification techniques generally achieve an AUC of around 87 %. As a by-product, a predictive inference on the risk of DR based on a patient's personalized medical information is derived. DiscussionMissing labels is a common issue for EHR data quality. Ignoring these missing labels can lead to biased results in the analyses of EHR data. The problem is especially severe in the context of DR. It is thus important to use machine learning or statistical tools to identify the unlabeled patients. The tool in this paper helps both data analysts and clinicians in their practices.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Estimating the prevalence of diabetic retinopathy in electronic health records with massive missing labels

Abstract

Talk to us

Similar Papers

More From: Intelligence-Based Medicine

Lead the way for us

Journal: Intelligence-Based Medicine	Publication Date: Jan 1, 2024
License type: cc-by-nc-nd

Similar Papers

Prevalence of self-reported diabetes and diabetic retinopathy in indigenous Australians: the National Indigenous Eye Health Survey
Jing Xie ... Hugh R Taylor
Clinical & Experimental Ophthalmology | VOL. 39
Jing Xie, et. al.Jing Xie ... Hugh R Taylor
24 Mar 2011
Clinical & Experimental Ophthalmology | VOL. 39

Prevalence of Diabetic Retinopathy in India: Sankara Nethralaya Diabetic Retinopathy Epidemiology and Molecular Genetics Study Report 2
Rajiv Raman ... Tarun Sharma
Ophthalmology | VOL. 116
Rajiv Raman, et. al.Rajiv Raman ... Tarun Sharma
12 Dec 2008
Ophthalmology | VOL. 116

From A Population to Patients: The Wisconsin Epidemiologic Study of Diabetic Retinopathy
Rohit Varma
Ophthalmology | VOL. 115
Rohit VarmaRohit Varma
01 Nov 2008
Ophthalmology | VOL. 115

Regional differences in the prevalence of diabetic retinopathy: a multi center study in Brazil
... Felipe Mallmann
Diabetology & Metabolic Syndrome | VOL. 10
, et. al. ... Felipe Mallmann
14 Mar 2018
Diabetology & Metabolic Syndrome | VOL. 10

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Estimating the prevalence of diabetic retinopathy in electronic health records with massive missing labels

Abstract

Talk to us

Similar Papers

More From: Intelligence-Based Medicine