Impact of Training Data, Ground Truth and Shape Variability in the Deep Learning-Based Semantic Segmentation of HeLa Cells Observed with Electron Microscopy-Reference-Cited by-同舟云学术

Impact of Training Data, Ground Truth and Shape Variability in the Deep Learning-Based Semantic Segmentation of HeLa Cells Observed with Electron Microscopy

Published:2023-03-01 Issue:3 Volume:9 Page:59
ISSN:2313-433X
Container-title:Journal of Imaging
language:en
Short-container-title:J. Imaging

Author:

Karabağ Cefa¹^ORCID,Ortega-Ruíz Mauricio Alberto¹²^ORCID,Reyes-Aldasoro Constantino Carlos¹^ORCID

Affiliation:

1. giCentre, Department of Computer Science, School of Science and Technology, City, University of London, London EC1V 0HB, UK

2. Departamento de Ingeniería, Campus Coyoacán, Universidad del Valle de México, Ciudad de México C.P. 04910, Mexico

Abstract

This paper investigates the impact of the amount of training data and the shape variability on the segmentation provided by the deep learning architecture U-Net. Further, the correctness of ground truth (GT) was also evaluated. The input data consisted of a three-dimensional set of images of HeLa cells observed with an electron microscope with dimensions 8192×8192×517. From there, a smaller region of interest (ROI) of 2000×2000×300 was cropped and manually delineated to obtain the ground truth necessary for a quantitative evaluation. A qualitative evaluation was performed on the 8192×8192 slices due to the lack of ground truth. Pairs of patches of data and labels for the classes nucleus, nuclear envelope, cell and background were generated to train U-Net architectures from scratch. Several training strategies were followed, and the results were compared against a traditional image processing algorithm. The correctness of GT, that is, the inclusion of one or more nuclei within the region of interest was also evaluated. The impact of the extent of training data was evaluated by comparing results from 36,000 pairs of data and label patches extracted from the odd slices in the central region, to 135,000 patches obtained from every other slice in the set. Then, 135,000 patches from several cells from the 8192×8192 slices were generated automatically using the image processing algorithm. Finally, the two sets of 135,000 pairs were combined to train once more with 270,000 pairs. As would be expected, the accuracy and Jaccard similarity index improved as the number of pairs increased for the ROI. This was also observed qualitatively for the 8192×8192 slices. When the 8192×8192 slices were segmented with U-Nets trained with 135,000 pairs, the architecture trained with automatically generated pairs provided better results than the architecture trained with the pairs from the manually segmented ground truths. This suggests that the pairs that were extracted automatically from many cells provided a better representation of the four classes of the various cells in the 8192×8192 slice than those pairs that were manually segmented from a single cell. Finally, the two sets of 135,000 pairs were combined, and the U-Net trained with these provided the best results.

Publisher

MDPI AG

Subject

Electrical and Electronic Engineering,Computer Graphics and Computer-Aided Design,Computer Vision and Pattern Recognition,Radiology, Nuclear Medicine and imaging

Link

https://www.mdpi.com/2313-433X/9/3/59/pdf

Reference96 articles.

1. HeLa cells 50 years on: The good, the bad and the ugly;Masters;Nat. Rev. Cancer,2002

2. A novel L1 retrotransposon marker for HeLa cell line identification;Rahbari;BioTechniques,2009

3. Identification of high-density lipoprotein in serum to determine anti-cancer efficacy of doxorubicin in HeLa cells;Yung;Int. J. Cancer,1992

4. Isolation and characterization of cancer stem cells from cervical cancer HeLa cells;Zhang;Cytotechnology,2012

5. Efflux excretion of bisdemethoxycurcumin-O-glucuronide in UGT1A1-overexpressing HeLa cells: Identification of breast cancer resistance protein (BCRP) and multidrug resistance-associated proteins 1 (MRP1) as the glucuronide transporters;Yang;Biofactors,2018

Cited by 6 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Dynamic label-free analysis of SARS-CoV-2 infection reveals virus-induced subcellular remodeling;Nature Communications;2024-06-11

2. Accurate detection of cell deformability tracking in hydrodynamic flow by coupling unsupervised and supervised learning;Machine Learning with Applications;2024-06

3. Fat-U-Net: non-contracting U-Net for free-space optical neural networks;AI and Optical Data Sciences V;2024-03-13

4. Dynamic label-free analysis of SARS-CoV-2 infection reveals virus-induced subcellular remodeling;2023-11-17

5. Opportunities and challenges for deep learning in cell dynamics research;Trends in Cell Biology;2023-11