Predictive Modeling of the Hospital Readmission Risk from Patients’ Claims Data Using Machine Learning: A Case Study on COPD

被引:0
作者
Xu Min
Bin Yu
Fei Wang
机构
[1] Weill Cornell Medicine,Department of Healthcare Policy and Research
[2] Tsinghua University,Department of Computer Science and Technology, Institute for Artificial Intelligence, Tsinghua
[3] American Air Liquide,Fuzhou Institute for Data Technology, and Bioinformatics Division, BNRist
来源
Scientific Reports | / 9卷
关键词
D O I
暂无
中图分类号
学科分类号
摘要
Chronic Obstructive Pulmonary Disease (COPD) is a prevalent chronic pulmonary condition that affects hundreds of millions of people all over the world. Many COPD patients got readmitted to hospital within 30 days after discharge due to various reasons. Such readmission can usually be avoided if additional attention is paid to patients with high readmission risk and appropriate actions are taken. This makes early prediction of the hospital readmission risk an important problem. The goal of this paper is to conduct a systematic study on developing different types of machine learning models, including both deep and non-deep ones, for predicting the readmission risk of COPD patients. We evaluate those different approaches on a real world database containing the medical claims of 111,992 patients from the Geisinger Health System from January 2004 to September 2015. The patient features we build the machine learning models upon include both knowledge-driven ones, which are the features extracted according to clinical knowledge potentially related to COPD readmission, and data-driven features, which are extracted from the patient data themselves. Our analysis showed that the prediction performance in terms of Area Under the receiver operating characteristic (ROC) Curve (AUC) can be improved from around 0.60 using knowledge-driven features, to 0.653 by combining both knowledge-driven and data-driven features, based on the one-year claims history before discharge. Moreover, we also demonstrate that the complex deep learning models in this case cannot really improve the prediction performance, with the best AUC around 0.65.
引用
收藏
相关论文
共 35 条
[1]  
Purdy S(2010)Prioritizing ambulatory care sensitive hospital admissions in england for research and intervention: a delphi exercise Prim. Heal. Care Res. & Dev. 11 41-50
[2]  
Griffin T(2017)Hospital readmissions for copd: a retrospective longitudinal study NPJ primary care respiratory medicine 27 100-105
[3]  
Salisbury C(2003)Risk factors of readmission to hospital for a copd exacerbation: a prospective study Thorax 58 551-557
[4]  
Sharp D(2010)Derivation and validation of an index to predict early death or unplanned readmission after discharge from hospital to the community Can. Med. Assoc. J. 182 632-638
[5]  
Harries TH(2013)Potentially avoidable 30-day hospital readmissions in medical patients: derivation and validation of a prediction model JAMA internal medicine 173 436-84
[6]  
Garcia-Aymerich J(2015)Deep learning nature 521 e0195024-365
[7]  
vanWalraven C(2018)Readmission prediction via deep contextual embedding of clinical concepts PloS one 13 77-297
[8]  
Donzé J(2012)Probabilistic topic models Commun. ACM 55 18-408
[9]  
Aujesky D(2018)Scalable and accurate deep learning with electronic health records. NPJ Digit Medicine 1 347-32
[10]  
Williams D(2013)Global strategy for the diagnosis, management, and prevention of chronic obstructive pulmonary disease: Gold executive summary Am. journal respiratory critical care medicine 187 273-1780