Using Generative AI to Improve the Performance and Interpretability of Rule-Based Diagnosis of Type 2 Diabetes Mellitus-Reference-Cited by-同舟云学术

Using Generative AI to Improve the Performance and Interpretability of Rule-Based Diagnosis of Type 2 Diabetes Mellitus

Published:2024-03-12 Issue:3 Volume:15 Page:162
ISSN:2078-2489
Container-title:Information
language:en
Short-container-title:Information

Author:

Kopitar Leon¹²^ORCID,Fister Iztok²^ORCID,Stiglic Gregor¹²³^ORCID

Affiliation:

1. Faculty of Health Sciences, University of Maribor, Zitna Ulica 15, 2000 Maribor, Slovenia

2. Faculty of Electrical Engineering and Computer Science, University of Maribor, Koroska Cesta 46, 2000 Maribor, Slovenia

3. Usher Institute, University of Edinburgh, Teviot Place, Edinburgh 0131, UK

Abstract

Introduction: Type 2 diabetes mellitus is a major global health concern, but interpreting machine learning models for diagnosis remains challenging. This study investigates combining association rule mining with advanced natural language processing to improve both diagnostic accuracy and interpretability. This novel approach has not been explored before in using pretrained transformers for diabetes classification on tabular data. Methods: The study used the Pima Indians Diabetes dataset to investigate Type 2 diabetes mellitus. Python and Jupyter Notebook were employed for analysis, with the NiaARM framework for association rule mining. LightGBM and the dalex package were used for performance comparison and feature importance analysis, respectively. SHAP was used for local interpretability. OpenAI GPT version 3.5 was utilized for outcome prediction and interpretation. The source code is available on GitHub. Results: NiaARM generated 350 rules to predict diabetes. LightGBM performed better than the GPT-based model. A comparison of GPT and NiaARM rules showed disparities, prompting a similarity score analysis. LightGBM’s decision making leaned heavily on glucose, age, and BMI, as highlighted in feature importance rankings. Beeswarm plots demonstrated how feature values correlate with their influence on diagnosis outcomes. Discussion: Combining association rule mining with GPT for Type 2 diabetes mellitus classification yields limited effectiveness. Enhancements like preprocessing and hyperparameter tuning are required. Interpretation challenges and GPT’s dependency on provided rules indicate the necessity for prompt engineering and similarity score methods. Variations in feature importance rankings underscore the complexity of T2DM. Concerns regarding GPT’s reliability emphasize the importance of iterative approaches for improving prediction accuracy.

Funder

Slovenian Research Agency

Publisher

MDPI AG

Link

https://www.mdpi.com/2078-2489/15/3/162/pdf

Reference48 articles.

1. IDF Diabetes Atlas: Global, regional and country-level diabetes prevalence estimates for 2021 and projections for 2045;Sun;Diabetes Res. Clin. Pract.,2022

2. Oh, S.H., Lee, S.J., and Park, J. (2022). Precision medicine for hypertension patients with type 2 diabetes via reinforcement learning. J. Pers. Med., 12.

3. Diabetic retinopathy screening using artificial intelligence and handheld smartphone-based retinal camera;Malerbi;J. Diabetes Sci. Technol.,2022

4. A deep learning nomogram of continuous glucose monitoring data for the risk prediction of diabetic retinopathy in type 2 diabetes;Tao;Phys. Eng. Sci. Med.,2023

5. ChatGPT in medicine: An overview of its applications, advantages, limitations, future prospects, and ethical considerations;Dave;Front. Artif. Intell.,2023

Cited by 1 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Generative Large Language Models in Electronic Health Records for Patient Care Since 2023: A Systematic Review;2024-08-12