Improving Multi-Tumor Biomarker Health Check-Up Tests with Machine Learning Algorithms-Reference-Cited by-同舟云学术

Improving Multi-Tumor Biomarker Health Check-Up Tests with Machine Learning Algorithms

Published:2020-06-01 Issue:6 Volume:12 Page:1442
ISSN:2072-6694
Container-title:Cancers
language:en
Short-container-title:Cancers

Author:

Wang Hsin-Yao^ORCID,Chen Chun-Hsien,Shi Steve,Chung Chia-Ru^ORCID,Wen Ying-Hao,Wu Min-Hsien,Lebowitz Michael S.,Zhou Jiming,Lu Jang-Jih

Abstract

Background: Tumor markers are used to screen tens of millions of individuals worldwide at annual health check-ups, especially in East Asia. Machine learning (ML)-based algorithms that improve the diagnostic accuracy and clinical utility of these tests can have substantial impact leading to the early diagnosis of cancer. Methods: ML-based algorithms, including a cancer screening algorithm and a secondary organ of origin algorithm, were developed and validated using a large real world dataset (RWD) from asymptomatic individuals undergoing routine cancer screening at a Taiwanese medical center between May 2001 and April 2015. External validation was performed using data from the same period from a separate medical center. The data set included tumor marker values, age, and gender from 27,938 individuals, including 342 subsequently confirmed cancer cases. Results: Separate gender-specific cancer screening algorithms were developed. For men, a logistic regression-based algorithm outperformed single-marker and other ML-based algorithms, with a mean area under the receiver operating characteristic curve (AUROC) of 0.7654 in internal and 0.8736 in external cross validation. For women, a random forest-based algorithm attained a mean AUROC of 0.6665 in internal and 0.6938 in external cross validation. The median time to cancer diagnosis (TTD) in men was 451.5, 204.5, and 28 days for the mild, moderate, and high-risk groups, respectively; for women, the median TTD was 229, 132, and 125 days for the mild, moderate, and high-risk groups. A second algorithm was developed to predict the most likely affected organ systems for at-risk individuals. The algorithm yielded 0.8120 sensitivity and 0.6490 specificity for men, and 0.8170 sensitivity and 0.6750 specificity for women. Conclusions: ML-derived algorithms, trained and validated by using a RWD, can significantly improve tumor marker-based screening for multiple types of early stage cancers, suggest the tissue of origin, and provide guidance for patient follow-up.

Publisher

MDPI AG

Subject

Cancer Research,Oncology

Link

https://www.mdpi.com/2072-6694/12/6/1442/pdf

Reference39 articles.

1. Cancer statistics, 2017

2. The Path to Cancer — Three Strikes and You're Out

3. Cancers Screening in an Asymptomatic Population by Using Multiple Tumour Markers

4. Assessment of quality in screening colonoscopy for colorectal cancer

5. Breast ultrasound: recommendations for information to women and referring physicians by the European Society of Breast Imaging

Cited by 17 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Identifying heterogeneous subgroups of systemic autoimmune diseases by applying a joint dimension reduction and clustering approach to immunomarkers;2024-04-30

2. Integrating Artificial Intelligence for Advancing Multiple-Cancer Early Detection via Serum Biomarkers: A Narrative Review;Cancers;2024-02-21

3. Multi-Cancer Early Detection: The New Frontier in Cancer Early Detection;Annual Review of Medicine;2024-01-29

4. Gender-Specific Machine Learning Models to Predict Unplanned Return to Operating Room Following Primary Total Shoulder Arthroplasty;2023 IEEE 11th International Conference on Healthcare Informatics (ICHI);2023-06-26

5. The Diagnostic Power of Circulating miR-1246 in Screening Cancer: An Updated Meta-analysis;Oxidative Medicine and Cellular Longevity;2023-04-20