Deconfounding Actor-Critic Network with Policy Adaptation for Dynamic Treatment Regimes.

Changchang Yin,Jeffrey Caterino,Ping Zhang,Ruoqi Liu

doi:10.1145/3534678.3539413

Abstract

Despite intense efforts in basic and clinical research, an individualized ventilation strategy for critically ill patients remains a major challenge. Recently, dynamic treatment regime (DTR) with reinforcement learning (RL) on electronic health records (EHR) has attracted interest from both the healthcare industry and machine learning research community. However, most learned DTR policies might be biased due to the existence of confounders. Although some treatment actions non-survivors received may be helpful, if confounders cause the mortality, the training of RL models guided by long-term outcomes (e.g., 90-day mortality) would punish those treatment actions causing the learned DTR policies to be suboptimal. In this study, we develop a new deconfounding actor-critic network (DAC) to learn optimal DTR policies for patients. To alleviate confounding issues, we incorporate a patient resampling module and a confounding balance module into our actor-critic framework. To avoid punishing the effective treatment actions non-survivors received, we design a short-term reward to capture patients' immediate health state changes. Combining short-term with long-term rewards could further improve the model performance. Moreover, we introduce a policy adaptation method to successfully transfer the learned model to new-source small-scale datasets. The experimental results on one semi-synthetic and two different real-world datasets show the proposed model outperforms the state-of-the-art models. The proposed model provides individualized treatment decisions for mechanical ventilation that could improve patient outcomes.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Deconfounding Actor-Critic Network with Policy Adaptation for Dynamic Treatment Regimes.

Abstract

Talk to us

Similar Papers

More From: KDD : proceedings. International Conference on Knowledge Discovery & Data Mining

Lead the way for us

Journal: KDD : proceedings. International Conference on Knowledge Discovery & Data Mining	Publication Date: Aug 14, 2022
Citations: 3

Similar Papers

Reinforcement Learning for Clinical Applications.
Kia Khezeli ... Benjamin Shickel
Clinical journal of the American Society of Nephrology : CJASN | VOL. 18
Kia Khezeli, et. al.Kia Khezeli ... Benjamin Shickel
08 Feb 2023
Clinical journal of the American Society of Nephrology : CJASN | VOL. 18

“The trouble with QALYs…”
Martin Knapp ... Roshni Mangalore
Epidemiology and Psychiatric Sciences | VOL. 16
Martin Knapp, et. al.Martin Knapp ... Roshni Mangalore
01 Dec 2007
Epidemiology and Psychiatric Sciences | VOL. 16

Projection based inverse reinforcement learning for the analysis of dynamic treatment regimes
Syed Ihtesham Hussain Shah ... Antonio Coronato
Applied Intelligence | VOL. 53
Syed Ihtesham Hussain Shah, et. al.Syed Ihtesham Hussain Shah ... Antonio Coronato
21 Oct 2022
Applied Intelligence | VOL. 53

ADT2R: Adaptive Decision Transformer for Dynamic Treatment Regimes in Sepsis.
Eunjin Jeon ... Heung-Il Suk
IEEE transactions on neural networks and learning systems | VOL. PP
Eunjin Jeon, et. al.Eunjin Jeon ... Heung-Il Suk
29 Aug 2024
IEEE transactions on neural networks and learning systems | VOL. PP

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Deconfounding Actor-Critic Network with Policy Adaptation for Dynamic Treatment Regimes.

Abstract

Talk to us

Similar Papers

More From: KDD : proceedings. International Conference on Knowledge Discovery & Data Mining