Why the study?
Most value-based deep reinforcement learning algorithms cannot precisely evaluate the target value function and are not as safe as clinical experts when assisting clinicians in sepsis treatment.
A novel deep reinforcement learning model with embedded human expertise achieved a high simulated survival rate for sepsis treatment.
Should not change sepsis care; leaves open whether WD3QNE improves outcomes in prospective trials.
Deep Reinforcement Learning (DRL) has been increasingly attempted in assisting clinicians for real-time treatment of sepsis. While a value function quantifies the performance of policies in such decision-making processes, most value-based DRL algorithms cannot evaluate the target value function precisely and are not as safe as clinical experts. In this study, we propose a Weighted Dueling Double Deep Q-Network with embedded human Expertise (WD3QNE). A target Q value function with adaptive dynamic weight is designed to improve the estimate accuracy and human expertise in decision-making is leveraged. In addition, the random forest algorithm is employed for feature selection to improve model interpretability. We test our algorithm against state-of-the-art value function methods in terms of expected return, survival rate, action distribution and external validation. The results demonstrate that WD3QNE obtains the highest survival rate of 97.81% in MIMIC-III dataset. Our proposed method is capable of providing reliable treatment decisions with embedded clinician expertise.
No takes yet. Share an insight, caveat, or question.
Wu et al. (2023) studied this question.
Synapse has enriched 2 closely related papers on similar clinical questions. Consider them for comparative context: