Modeling the Risk of Liver Cancer in Adults: A Machine Learning Investigation into the Role of Obesity and Overweight

Author:

Karamo Bah,Adama Ns Bah,Amadou Wurry Jallow

Abstract

Background: Liver cancer is a global health concern, with overweight and obese individuals exhibiting an increased risk of its development. Understanding the interplay between obesity-related factors and liver cancer incidence is crucial for early prediction and intervention. Aim: The aim of this investigation was to construct and validate an extreme gradient boosting (XGBoost) based machine learning model for the purpose of establishing a one-year liver cancer risk prediction system specifically tailored to overweight and obese patients. In addition, this study sought to compare the predictive performance of the XGBoost model with those of a random forest model and a logistic regression model, while also identifying the most influential predictive features for liver cancer incidence. Methods: A comprehensive retrospective analysis was conducted on MIMIC III data comprising 2,354 patients. To predict the risk of liver cancer development, three machine learning models were developed: XGBoost, random forest, and logistic regression. Feature selection was executed using a stepwise regression procedure encompassing both forward selection and backward elimination. Results: The stepwise regression technique unveiled 14 predictive factors for liver cancer incidence. Among the patient cohort, 132 individuals developed liver cancer within a year of follow-up, while 2,222 did not. Notably, most liver cancer cases occurred in male patients (60%). Statistically significant differences were observed between patients with liver cancer and those without, in terms of age, gender, total bilirubin, platelet, albumin, chloride, potassium, sodium, prothrombin time (PT) and alanine aminotransferase (ALT). The XGBoost model exhibited an impressive area under the receiver operating characteristic curve (AUROC) of 99%, Random Forest (RF) of 99%, and Logistic Regression (LR) of 90%. In a multivariate analysis, total bilirubin, creatinine levels, age, gender, ALT, alkaline phosphate (ALP), PT, calcium, and chloride emerged as independent predictors for liver cancer incidence. Conclusion: The XGBoost model demonstrated superior predictive performance when compared to the RF and LR models. If corroborated through prospective studies, the XGBoost model may prove to be a valuable tool for the early prediction of liver cancer risk in overweight or obese individuals. Such predictive capabilities could, in turn, facilitate the implementation of timely preventive interventions against liver cancer.

Publisher

Heighten Science Publications Corporation

Subject

Materials Chemistry

同舟云学术

1.学者识别学者识别

2.学术分析学术分析

3.人才评估人才评估

"同舟云学术"是以全球学者为主线,采集、加工和组织学术论文而形成的新型学术文献查询和分析系统,可以对全球学者进行文献检索和人才价值评估。用户可以通过关注某些学科领域的顶尖人物而持续追踪该领域的学科进展和研究前沿。经过近期的数据扩容,当前同舟云学术共收录了国内外主流学术期刊6万余种,收集的期刊论文及会议论文总量共计约1.5亿篇,并以每天添加12000余篇中外论文的速度递增。我们也可以为用户提供个性化、定制化的学者数据。欢迎来电咨询!咨询电话:010-8811{复制后删除}0370

www.globalauthorid.com

TOP

Copyright © 2019-2024 北京同舟云网络信息技术有限公司
京公网安备11010802033243号  京ICP备18003416号-3