Using Automated Machine Learning to Predict the Mortality of Patients With COVID-19: Prediction Model Development Study-Reference-Cited by-同舟云学术

Using Automated Machine Learning to Predict the Mortality of Patients With COVID-19: Prediction Model Development Study

Published:2021-02-26 Issue:2 Volume:23 Page:e23458
ISSN:1438-8871
Container-title:Journal of Medical Internet Research
language:en
Short-container-title:J Med Internet Res

Author:

Ikemura Kenji^ORCID,Bellin Eran^ORCID,Yagi Yukako^ORCID,Billett Henny^ORCID,Saada Mahmoud^ORCID,Simone Katelyn^ORCID,Stahl Lindsay^ORCID,Szymanski James^ORCID,Goldstein D Y^ORCID,Reyes Gil Morayma^ORCID

Abstract

Background During a pandemic, it is important for clinicians to stratify patients and decide who receives limited medical resources. Machine learning models have been proposed to accurately predict COVID-19 disease severity. Previous studies have typically tested only one machine learning algorithm and limited performance evaluation to area under the curve analysis. To obtain the best results possible, it may be important to test different machine learning algorithms to find the best prediction model. Objective In this study, we aimed to use automated machine learning (autoML) to train various machine learning algorithms. We selected the model that best predicted patients’ chances of surviving a SARS-CoV-2 infection. In addition, we identified which variables (ie, vital signs, biomarkers, comorbidities, etc) were the most influential in generating an accurate model. Methods Data were retrospectively collected from all patients who tested positive for COVID-19 at our institution between March 1 and July 3, 2020. We collected 48 variables from each patient within 36 hours before or after the index time (ie, real-time polymerase chain reaction positivity). Patients were followed for 30 days or until death. Patients’ data were used to build 20 machine learning models with various algorithms via autoML. The performance of machine learning models was measured by analyzing the area under the precision-recall curve (AUPCR). Subsequently, we established model interpretability via Shapley additive explanation and partial dependence plots to identify and rank variables that drove model predictions. Afterward, we conducted dimensionality reduction to extract the 10 most influential variables. AutoML models were retrained by only using these 10 variables, and the output models were evaluated against the model that used 48 variables. Results Data from 4313 patients were used to develop the models. The best model that was generated by using autoML and 48 variables was the stacked ensemble model (AUPRC=0.807). The two best independent models were the gradient boost machine and extreme gradient boost models, which had an AUPRC of 0.803 and 0.793, respectively. The deep learning model (AUPRC=0.73) was substantially inferior to the other models. The 10 most influential variables for generating high-performing models were systolic and diastolic blood pressure, age, pulse oximetry level, blood urea nitrogen level, lactate dehydrogenase level, D-dimer level, troponin level, respiratory rate, and Charlson comorbidity score. After the autoML models were retrained with these 10 variables, the stacked ensemble model still had the best performance (AUPRC=0.791). Conclusions We used autoML to develop high-performing models that predicted the survival of patients with COVID-19. In addition, we identified important variables that correlated with mortality. This is proof of concept that autoML is an efficient, effective, and informative method for generating machine learning–based clinical decision support tools.

Publisher

JMIR Publications Inc.

Subject

Health Informatics

Reference29 articles.

1. COVID data trackerCenters for Disease Control and Prevention2021-02-12https://www.cdc.gov/coronavirus/2019-ncov/cases-updates/cases-in-us.html

2. Using Machine Learning to Predict ICU Transfer in Hospitalized COVID-19 Patients

3. Machine Learning to Predict Mortality and Critical Events in a Cohort of Patients With COVID-19 in New York City: Model Development and Validation

4. An interpretable mortality prediction model for COVID-19 patients

5. Montefiore Medical CenterTsubomi Technology2021-02-23https://www.tsubomi.tech/

Cited by 60 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Automated machine learning with interpretation: A systematic review of methodologies and applications in healthcare;Medicine Advances;2024-08-27

2. Respiratory Tract Pathogen Profiles of COVID-19 Pneumonia Patients and the Mortality Prediction;2024-07-22

3. At-admission prediction of mortality and pulmonary embolism in an international cohort of hospitalised patients with COVID-19 using statistical and machine learning methods;Scientific Reports;2024-07-16

4. Artificial Intelligence in Medicine: from Diagnosis to Treatment;Futurity Medicine;2024-07-10

5. Explainable coronary artery disease prediction model based on AutoGluon from AutoML framework;Frontiers in Cardiovascular Medicine;2024-07-01