Abstract
AbstractIn this study, we leveraged machine learning (ML) approach to develop and validate new assessment tools for predicting stroke and bleeding among patients with atrial fibrillation (AFib) and cancer. We conducted a retrospective cohort study including patients who were newly diagnosed with AFib with a record of cancer from the 2012–2018 Surveillance, Epidemiology, and End Results (SEER)-Medicare database. The ML algorithms were developed and validated separately for each outcome by fitting elastic net, random forest (RF), extreme gradient boosting (XGBoost), support vector machine (SVM), and neural network models with tenfold cross-validation (train:test = 7:3). We obtained area under the curve (AUC), sensitivity, specificity, and F2 score as performance metrics. Model calibration was assessed using Brier score. In sensitivity analysis, we resampled data using Synthetic Minority Oversampling Technique (SMOTE). Among 18,388 patients with AFib and cancer, 523 (2.84%) had ischemic stroke and 221 (1.20%) had major bleeding within one year after AFib diagnosis. In prediction of ischemic stroke, RF significantly outperformed other ML models [AUC (0.916, 95% CI 0.887–0.945), sensitivity 0.868, specificity 0.801, F2 score 0.375, Brier score = 0.035]. However, the performance of ML algorithms in prediction of major bleeding was low with highest AUC achieved by RF (0.623, 95% CI 0.554–0.692). RF models performed better than CHA2DS2-VASc and HAS-BLED scores. SMOTE did not improve the performance of the ML algorithms. Our study demonstrated a promising application of ML in stroke prediction among patients with AFib and cancer. This tool may be leveraged in assisting clinicians to identify patients at high risk of stroke and optimize treatment decisions.
Publisher
Springer Science and Business Media LLC
Reference63 articles.
1. Patel, N. J., Deshmukh, A., Pant, S., et al. (2014). Contemporary trends of hospitalization for atrial fibrillation in the United States, 2000 through 2010: Implications for healthcare planning. Circulation, 129(23), 2371–2379.
2. Benjamin, E. J., Muntner, P., Alonso, A., et al. (2019). Heart disease and stroke statistics-2019 update: A report from the American Heart Association. Circulation, 139(10), e56–e528.
3. Centers for Disease Control and Prevention - National Center for Health Statistics. About Multiple Cause of Death, 1999–2019. https://wonder.cdc.gov/mcd-icd10.html. Published 2019. Retrieved October 14, 2021.
4. Chung, M. K., Eckhardt, L. L., Chen, L. Y., et al. (2020). Lifestyle and risk factor modification for reduction of atrial fibrillation: A Scientific statement from the American Heart Association. Circulation, 141(16), e750–e772.
5. Timp, J. F., Braekkan, S. K., Versteeg, H. H., & Cannegieter, S. C. (2013). Epidemiology of cancer-associated venous thrombosis. Blood, 122(10), 1712–1723.