PATIENT RECRUITMENT USING ELECTRONIC HEALTH RECORDS UNDER SELECTION BIAS: A TWO-PHASE SAMPLING FRAMEWORK.

Guanghao Zhang,Lauren J Beesley,Bhramar Mukherjee,X U Shi

doi:10.1214/23-aoas1860

Abstract

Electronic health records (EHRs) are increasingly recognized as a cost-effective resource for patient recruitment in clinical research. However, how to optimally select a cohort from millions of individuals to answer a scientific question of interest remains unclear. Consider a study to estimate the mean or mean difference of an expensive outcome. Inexpensive auxiliary covariates predictive of the outcome may often be available in patients' health records, presenting an opportunity to recruit patients selectively, which may improve efficiency in downstream analyses. In this paper we propose a two-phase sampling design that leverages available information on auxiliary covariates in EHR data. A key challenge in using EHR data for multiphase sampling is the potential selection bias, because EHR data are not necessarily representative of the target population. Extending existing literature on two-phase sampling design, we derive an optimal two-phase sampling method that improves efficiency over random sampling while accounting for the potential selection bias in EHR data. We demonstrate the efficiency gain from our sampling design via simulation studies and an application evaluating the prevalence of hypertension among U.S. adults leveraging data from the Michigan Genomics Initiative, a longitudinal biorepository in Michigan Medicine.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

PATIENT RECRUITMENT USING ELECTRONIC HEALTH RECORDS UNDER SELECTION BIAS: A TWO-PHASE SAMPLING FRAMEWORK.

Abstract

Talk to us

Similar Papers

More From: The annals of applied statistics

Lead the way for us

Similar Papers

HIR Collaborating with the CODATA Conference
Hyejung Chang ... William T F Goossen
Healthcare Informatics Research | VOL. 19
Hyejung Chang, et. al.Hyejung Chang ... William T F Goossen
01 Jan 2013
Healthcare Informatics Research | VOL. 19

Continuity and Completeness of Electronic Health Record Data for Patients Treated With Oral Hypoglycemic Agents: Findings From Healthcare Delivery Systems in Taiwan.
Chien-Ning Hsu ... Kelly Huang
Frontiers in Pharmacology | VOL. 13
Chien-Ning Hsu, et. al.Chien-Ning Hsu ... Kelly Huang
04 Apr 2022
Frontiers in Pharmacology | VOL. 13

Enhancement in line of therapy (LoT) derivation from real-world data (RWD) from electronic health records (EHR) via integration of medical claims data.
Smita Agrawal ... Vandana Priya
Journal of Clinical Oncology | VOL. 41
Smita Agrawal, et. al.Smita Agrawal ... Vandana Priya
01 Jun 2023
Journal of Clinical Oncology | VOL. 41

Data extraction from electronic health records (EHRs) for quality measurement of the physical therapy process: comparison between EHR data and survey data.
Marijn Scholte ... Jozé Braspenning
BMC Medical Informatics and Decision Making | VOL. 16
Marijn Scholte, et. al.Marijn Scholte ... Jozé Braspenning
08 Nov 2016
BMC Medical Informatics and Decision Making | VOL. 16

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

PATIENT RECRUITMENT USING ELECTRONIC HEALTH RECORDS UNDER SELECTION BIAS: A TWO-PHASE SAMPLING FRAMEWORK.

Abstract

Talk to us

Similar Papers

More From: The annals of applied statistics