Construction of the prediction model for multiple myeloma based on machine learning-Reference-Cited by-同舟云学术

Construction of the prediction model for multiple myeloma based on machine learning

Published:2024-05-31 Issue: Volume: Page:
ISSN:1751-5521
Container-title:International Journal of Laboratory Hematology
language:en
Short-container-title:Int J Lab Hematology

Author:

Cai Jiangying¹^ORCID,Liu Zhenhua¹,Wang Yingying¹,Yang Wanxia¹,Sun Zhipeng²,You Chongge¹

Affiliation:

1. The Second Hospital & Clinical Medical School Lanzhou University Lanzhou People's Republic of China

2. Department of Scientific & Application Sysmex Shanghai Ltd Shanghai People's Republic of China

Abstract

AbstractIntroductionThe global burden of multiple myeloma (MM) is increasing every year. Here, we have developed machine learning models to provide a reference for the early detection of MM.MethodsA total of 465 patients and 150 healthy controls were enrolled in this retrospective study. Based on the variable screening strategy of least absolute shrinkage and selection operator (LASSO), three prediction models, logistic regression (LR), support vector machine (SVM), and random forest (RF), were established combining complete blood count (CBC) and cell population data (CPD) parameters in the training set (210 cases), and were verified in the validation set (90 cases) and test set (165 cases). The performance of each model was analyzed using receiver operating characteristic (ROC) curve, calibration curves, and decision curve analysis (DCA). Accuracy, sensitivity, specificity, positive predictive value, negative predictive value, and area under the ROC curve (AUC) were applied to evaluate the models. Delong test was used to compare the AUC of the models.ResultsSix parameters including RBC (1012/L), RDW‐CV (%), IG (%), NE‐WZ, LY‐WX, and LY‐WZ were screened out by LASSO to construct the model. Among the three models, the AUC of RF model in the training set, validation set, and test set were 0.956, 0.892, and 0.875, which were higher than those of LR model (0.901, 0.849, and 0.858) and SVM model (0.929, 0.868, and 0.846). Delong test showed that there were significant differences among the models in the training set, no significant differences in the validation set, and significant differences only between SVM and RF models in the test set. The calibration curve and DCA showed that the three models had good validity and feasibility, and the RF model performed best.ConclusionThe proposed RF model may be a useful auxiliary tool for rapid screening of MM patients.

Publisher

Wiley

Link

https://onlinelibrary.wiley.com/doi/pdf/10.1111/ijlh.14324

Reference33 articles.

1. Multiple Myeloma

2. Incidence and mortality of multiple myeloma in China, 2006–2016: an analysis of the Global Burden of Disease Study 2016

3. Mortality of lymphoma and myeloma in China, 2004–2017: an observational study

4. Global cancer statistics 2018: GLOBOCAN estimates of incidence and mortality worldwide for 36 cancers in 185 countries

5. Global Cancer Statistics 2020: GLOBOCAN Estimates of Incidence and Mortality Worldwide for 36 Cancers in 185 Countries