Dimension Reduction and Clustering Models for Single-Cell RNA Sequencing Data: A Comparative Study-Reference-Cited by-同舟云学术

Dimension Reduction and Clustering Models for Single-Cell RNA Sequencing Data: A Comparative Study

Published:2020-03-22 Issue:6 Volume:21 Page:2181
ISSN:1422-0067
Container-title:International Journal of Molecular Sciences
language:en
Short-container-title:IJMS

Author:

Feng Chao,Liu Shufen,Zhang Hao,Guan Renchu^ORCID,Li Dan,Zhou Fengfeng^ORCID,Liang Yanchun,Feng Xiaoyue^ORCID

Abstract

With recent advances in single-cell RNA sequencing, enormous transcriptome datasets have been generated. These datasets have furthered our understanding of cellular heterogeneity and its underlying mechanisms in homogeneous populations. Single-cell RNA sequencing (scRNA-seq) data clustering can group cells belonging to the same cell type based on patterns embedded in gene expression. However, scRNA-seq data are high-dimensional, noisy, and sparse, owing to the limitation of existing scRNA-seq technologies. Traditional clustering methods are not effective and efficient for high-dimensional and sparse matrix computations. Therefore, several dimension reduction methods have been introduced. To validate a reliable and standard research routine, we conducted a comprehensive review and evaluation of four classical dimension reduction methods and five clustering models. Four experiments were progressively performed on two large scRNA-seq datasets using 20 models. Results showed that the feature selection method contributed positively to high-dimensional and sparse scRNA-seq data. Moreover, feature-extraction methods were able to promote clustering performance, although this was not eternally immutable. Independent component analysis (ICA) performed well in those small compressed feature spaces, whereas principal component analysis was steadier than all the other feature-extraction methods. In addition, ICA was not ideal for fuzzy C-means clustering in scRNA-seq data analysis. K-means clustering was combined with feature-extraction methods to achieve good results.

Funder

the National Natural Science Foundation of China

Publisher

MDPI AG

Subject

Inorganic Chemistry,Organic Chemistry,Physical and Theoretical Chemistry,Computer Science Applications,Spectroscopy,Molecular Biology,General Medicine,Catalysis

Link

https://www.mdpi.com/1422-0067/21/6/2181/pdf

Reference69 articles.

1. Rare cell isolation and analysis in microfluidics

2. Landscape of Infiltrating T Cells in Liver Cancer Revealed by Single-Cell Sequencing

3. Global characterization of T cells in non-small-cell lung cancer by single-cell sequencing

4. Challenges in unsupervised clustering of single-cell RNA-seq data

5. The Human Cell Atlas

Cited by 39 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Interpreting single-cell and spatial omics data using deep networks training dynamics;2024-04-10

2. Clustering Analysis of Time Series of Affect in Dyadic Interactions;Multivariate Behavioral Research;2024-02-26

3. ClustML: A measure of cluster pattern complexity in scatterplots learnt from human-labeled groupings;Information Visualization;2024-01-30

4. Dimensionality Reduction and Clustering;SpringerBriefs in Applied Sciences and Technology;2024

5. Deep generative model deciphers derailed trajectories in acute myeloid leukemia;2023-11-15