Impact of the Covid-19 pandemic on the performance of machine learning algorithms for predicting perioperative mortality-Reference-Cited by-同舟云学术

Impact of the Covid-19 pandemic on the performance of machine learning algorithms for predicting perioperative mortality

Published:2023-04-12 Issue:1 Volume:23 Page:
ISSN:1472-6947
Container-title:BMC Medical Informatics and Decision Making
language:en
Short-container-title:BMC Med Inform Decis Mak

Author:

Andonov D. I.,Ulm B.,Graessner M.,Podtschaske A.,Blobner M.,Jungwirth B.,Kagerbauer S. M.

Abstract

Abstract Background Machine-learning models are susceptible to external influences which can result in performance deterioration. The aim of our study was to elucidate the impact of a sudden shift in covariates, like the one caused by the Covid-19 pandemic, on model performance. Methods After ethical approval and registration in Clinical Trials (NCT04092933, initial release 17/09/2019), we developed different models for the prediction of perioperative mortality based on preoperative data: one for the pre-pandemic data period until March 2020, one including data before the pandemic and from the first wave until May 2020, and one that covers the complete period before and during the pandemic until October 2021. We applied XGBoost as well as a Deep Learning neural network (DL). Performance metrics of each model during the different pandemic phases were determined, and XGBoost models were analysed for changes in feature importance. Results XGBoost and DL provided similar performance on the pre-pandemic data with respect to area under receiver operating characteristic (AUROC, 0.951 vs. 0.942) and area under precision-recall curve (AUPR, 0.144 vs. 0.187). Validation in patient cohorts of the different pandemic waves showed high fluctuations in performance from both AUROC and AUPR for DL, whereas the XGBoost models seemed more stable. Change in variable frequencies with onset of the pandemic were visible in age, ASA score, and the higher proportion of emergency operations, among others. Age consistently showed the highest information gain. Models based on pre-pandemic data performed worse during the first pandemic wave (AUROC 0.914 for XGBoost and DL) whereas models augmented with data from the first wave lacked performance after the first wave (AUROC 0.907 for XGBoost and 0.747 for DL). The deterioration was also visible in AUPR, which worsened by over 50% in both XGBoost and DL in the first phase after re-training. Conclusions A sudden shift in data impacts model performance. Re-training the model with updated data may cause degradation in predictive accuracy if the changes are only transient. Too early re-training should therefore be avoided, and close model surveillance is necessary.

Funder

German Federal Ministry for Economic Affairs and Energy

Universität Ulm

Publisher

Springer Science and Business Media LLC

Subject

Health Informatics,Health Policy,Computer Science Applications

Link

https://link.springer.com/content/pdf/10.1186/s12911-023-02151-1.pdf

Reference28 articles.

1. Turley L, Mahon J, Sheehan E. “Out of hours” orthopaedics in an Irish regional trauma unit and the impact of COVID-19. Ir J Med Sci. 2022. https://doi.org/10.1007/s11845-022-03135-2.

2. Duckworth C, Chmiel FP, Burns DK, Zlatev ZD, White NM, Daniels TWV, Kiuber M, Boniface MJ. Using explainable machine learning to characterise data drift and detect emergent health risks for emergency department admissions during COVID-19. Sci Rep. 2021;11(1):23017. https://doi.org/10.1038/s41598-021-02481-y.

3. Chi S, Tian Y, Wang F, Zhou T, Jin S, Li J. A novel lifelong machine learning-based method to eliminate calibration drift in clinical prediction models. Artif Intell Med. 2022;125:102256. https://doi.org/10.1016/j.artmed.2022.102256.

4. Celik B, Vanschoren J. Adaptation Strategies for Automated Machine Learning on Evolving Data. IEEE Trans Pattern Anal Mach Intell. 2021;43(9):3067–78. https://doi.org/10.1109/TPAMI.2021.3062900.

5. Das S: Best Practices for Dealing With Concept Drift [https://neptune.ai/blog/concept-drift-best-practices] last Accessed 14 Nov 2022

Cited by 2 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Towards Improved XAI-Based Epidemiological Research into the Next Potential Pandemic;Life;2024-06-21

2. Susceptibility of AutoML mortality prediction algorithms to model drift caused by the COVID pandemic;BMC Medical Informatics and Decision Making;2024-02-02