A Multimodality Machine Learning Approach to Differentiate Severe and Nonsevere COVID-19: Model Development and Validation-Reference-Cited by-同舟云学术

A Multimodality Machine Learning Approach to Differentiate Severe and Nonsevere COVID-19: Model Development and Validation

Published:2021-04-07 Issue:4 Volume:23 Page:e23948
ISSN:1438-8871
Container-title:Journal of Medical Internet Research
language:en
Short-container-title:J Med Internet Res

Author:

Chen Yuanfang^ORCID,Ouyang Liu^ORCID,Bao Forrest S^ORCID,Li Qian^ORCID,Han Lei^ORCID,Zhang Hengdong^ORCID,Zhu Baoli^ORCID,Ge Yaorong^ORCID,Robinson Patrick^ORCID,Xu Ming^ORCID,Liu Jie^ORCID,Chen Shi^ORCID

Abstract

Background Effectively and efficiently diagnosing patients who have COVID-19 with the accurate clinical type of the disease is essential to achieve optimal outcomes for the patients as well as to reduce the risk of overloading the health care system. Currently, severe and nonsevere COVID-19 types are differentiated by only a few features, which do not comprehensively characterize the complicated pathological, physiological, and immunological responses to SARS-CoV-2 infection in the different disease types. In addition, these type-defining features may not be readily testable at the time of diagnosis. Objective In this study, we aimed to use a machine learning approach to understand COVID-19 more comprehensively, accurately differentiate severe and nonsevere COVID-19 clinical types based on multiple medical features, and provide reliable predictions of the clinical type of the disease. Methods For this study, we recruited 214 confirmed patients with nonsevere COVID-19 and 148 patients with severe COVID-19. The clinical characteristics (26 features) and laboratory test results (26 features) upon admission were acquired as two input modalities. Exploratory analyses demonstrated that these features differed substantially between two clinical types. Machine learning random forest models based on all the features in each modality as well as on the top 5 features in each modality combined were developed and validated to differentiate COVID-19 clinical types. Results Using clinical and laboratory results independently as input, the random forest models achieved >90% and >95% predictive accuracy, respectively. The importance scores of the input features were further evaluated, and the top 5 features from each modality were identified (age, hypertension, cardiovascular disease, gender, and diabetes for the clinical features modality, and dimerized plasmin fragment D, high sensitivity troponin I, absolute neutrophil count, interleukin 6, and lactate dehydrogenase for the laboratory testing modality, in descending order). Using these top 10 multimodal features as the only input instead of all 52 features combined, the random forest model was able to achieve 97% predictive accuracy. Conclusions Our findings shed light on how the human body reacts to SARS-CoV-2 infection as a unit and provide insights on effectively evaluating the disease severity of patients with COVID-19 based on more common medical features when gold standard features are not available. We suggest that clinical information can be used as an initial screening tool for self-evaluation and triage, while laboratory test results should be applied when accuracy is the priority.

Publisher

JMIR Publications Inc.

Subject

Health Informatics

Reference47 articles.

1. Weekly epidemiological update - 12 January 2021World Health Organization202101122021-03-18https://www.who.int/publications/m/item/weekly-epidemiological-update---12-january-2021

2. Asymptomatic Transmission, the Achilles’ Heel of Current Strategies to Control Covid-19

3. Virological assessment of hospitalized patients with COVID-2019

4. Asymptomatic SARS-CoV-2 Infection in Household Contacts of a Healthcare Provider, Wuhan, China

5. Delivery of infection from asymptomatic carriers of COVID-19 in a familial cluster

Cited by 30 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Multi-modality approaches for medical support systems: A systematic review of the last decade;Information Fusion;2024-03

2. Machine learning versus multivariate logistic regression for predicting severe COVID‐19 in hospitalized children with Omicron variant infection;Journal of Medical Virology;2024-02

3. Derivation and external validation of a nomogram predicting the occurrence of severe illness among hospitalized coronavirus disease 2019 patients: a 2020 Chinese multicenter retrospective study;Journal of Thoracic Disease;2023-12

4. Severity prediction in COVID-19 patients using clinical markers and explainable artificial intelligence: A stacked ensemble machine learning approach;Intelligent Decision Technologies;2023-11-20

5. Using artificial intelligence algorithms to predict the overall survival of hemodialysis patients during the COVID-19 pandemic: A prospective cohort study;Journal of the Chinese Medical Association;2023-09-15