Using autoencoders as a weight initialization method on deep neural networks for disease detection-Reference-Cited by-同舟云学术

Using autoencoders as a weight initialization method on deep neural networks for disease detection

Published:2020-08 Issue:S5 Volume:20 Page:
ISSN:1472-6947
Container-title:BMC Medical Informatics and Decision Making
language:en
Short-container-title:BMC Med Inform Decis Mak

Author:

Ferreira Mafalda Falcão,Camacho Rui,Teixeira Luís F.

Abstract

Abstract Background As of today, cancer is still one of the most prevalent and high-mortality diseases, summing more than 9 million deaths in 2018. This has motivated researchers to study the application of machine learning-based solutions for cancer detection to accelerate its diagnosis and help its prevention. Among several approaches, one is to automatically classify tumor samples through their gene expression analysis. Methods In this work, we aim to distinguish five different types of cancer through RNA-Seq datasets: thyroid, skin, stomach, breast, and lung. To do so, we have adopted a previously described methodology, with which we compare the performance of 3 different autoencoders (AEs) used as a deep neural network weight initialization technique. Our experiments consist in assessing two different approaches when training the classification model — fixing the weights after pre-training the AEs, or allowing fine-tuning of the entire network — and two different strategies for embedding the AEs into the classification network, namely by only importing the encoding layers, or by inserting the complete AE. We then study how varying the number of layers in the first strategy, the AEs latent vector dimension, and the imputation technique in the data preprocessing step impacts the network’s overall classification performance. Finally, with the goal of assessing how well does this pipeline generalize, we apply the same methodology to two additional datasets that include features extracted from images of malaria thin blood smears, and breast masses cell nuclei. We also discard the possibility of overfitting by using held-out test sets in the images datasets. Results The methodology attained good overall results for both RNA-Seq and image extracted data. We outperformed the established baseline for all the considered datasets, achieving an average F1 score of 99.03, 89.95, and 98.84 and an MCC of 0.99, 0.84, and 0.98, for the RNA-Seq (when detecting thyroid cancer), the Malaria, and the Wisconsin Breast Cancer data, respectively. Conclusions We observed that the approach of fine-tuning the weights of the top layers imported from the AE reached higher results, for all the presented experiences, and all the considered datasets. We outperformed all the previous reported results when comparing to the established baselines.

Publisher

Springer Science and Business Media LLC

Subject

Health Informatics,Health Policy,Computer Science Applications

Link

https://link.springer.com/content/pdf/10.1186/s12911-020-01150-w.pdf

Reference47 articles.

1. World Health Organization (WHO). Cancer. 2018. https://www.who.int/health-topics/cancer. Accessed on 22 Nov 2019.

2. World Health Organization (WHO). Cancer Fact Sheet. 2018. https://www.who.int/en/news-room/fact-sheets/detail/cancer. Accessed on 22 Nov 22 2019.

3. BC Cancer. Change in 5-year survival rates by cancer type for adults in BC, 1997 - 2016. 2019. http://www.bccancer.bc.ca/statistics-and-reports-site/Documents/Five_Year_Survival_Change_Report_2016. _20190321.pdf. Accessed on 22 Nov 2019.

4. Kourou K, Exarchos TP, Exarchos KP, Karamouzis MV, Fotiadis DI. Machine learning applications in cancer prognosis and prediction. Comput Struct Biotechnol J. 2015; 13:8–17.

5. National Center for Biotechnology Information (NCBI). Gene Expression. 2017. https://www.ncbi.nlm.nih.gov/probe/docs/applexpression/. Accessed on 25 Nov 2019.

Cited by 21 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. BCED-Net: Breast Cancer Ensemble Diagnosis Network using transfer learning and the XGBoost classifier with mammography images;Osong Public Health and Research Perspectives;2024-09-10

2. Computer-Aided Diagnosis Systems for Automatic Malaria Parasite Detection and Classification: A Systematic Review;Electronics;2024-08-11

3. The pre-trained explainable deep learning model with stacked denoising autoencoders for slope stability analysis;Engineering Analysis with Boundary Elements;2024-06

4. Deep Learning Methods for Diagnosing Thyroid Cancer;Journal of Engineering and Science in Medical Diagnostics and Therapy;2024-02-28

5. Improving Breast Cancer Prognosis with DL-Based Image Classification;Lecture Notes in Networks and Systems;2024