An integrated single-cell transcriptomic dataset for non-small cell lung cancer-Reference-Cited by-同舟云学术

An integrated single-cell transcriptomic dataset for non-small cell lung cancer

Published:2023-03-27 Issue:1 Volume:10 Page:
ISSN:2052-4463
Container-title:Scientific Data
language:en
Short-container-title:Sci Data

Author:

Prazanowska Karolina Hanna^ORCID,Lim Su Bin^ORCID

Abstract

AbstractAs single-cell RNA sequencing (scRNA-seq) has emerged as a great tool for studying cellular heterogeneity within the past decade, the number of available scRNA-seq datasets also rapidly increased. However, reuse of such data is often problematic due to a small cohort size, limited cell types, and insufficient information on cell type classification. Here, we present a large integrated scRNA-seq dataset containing 224,611 cells from human primary non-small cell lung cancer (NSCLC) tumors. Using publicly available resources, we pre-processed and integrated seven independent scRNA-seq datasets using an anchor-based approach, with five datasets utilized as reference and the remaining two, as validation. We created two levels of annotation based on cell type-specific markers conserved across the datasets. To demonstrate usability of the integrated dataset, we created annotation predictions for the two validation datasets using our integrated reference. Additionally, we conducted a trajectory analysis on subsets of T cells and lung cancer cells. This integrated data may serve as a resource for studying NSCLC transcriptome at the single cell level.

Funder

National Research Foundation of Korea

Ministry of Health and Welfare

Publisher

Springer Science and Business Media LLC

Subject

Library and Information Sciences,Statistics, Probability and Uncertainty,Computer Science Applications,Education,Information Systems,Statistics and Probability

Link

https://www.nature.com/articles/s41597-023-02074-6.pdf

Reference79 articles.

1. Tang, F. et al. mRNA-Seq whole-transcriptome analysis of a single cell. Nat Methods 6, 377–382, https://doi.org/10.1038/nmeth.1315 (2009).

2. Zhang, Y. et al. Single-cell RNA sequencing in cancer research. J Exp Clin Cancer Res 40, 81, https://doi.org/10.1186/s13046-021-01874-1 (2021).

3. Seow, J. J. W., Wong, R. M. M., Pai, R. & Sharma, A. Single-Cell RNA Sequencing for Precision Oncology: Current State-of-Art. J Indian Inst Sci 100, 579–588, https://doi.org/10.1007/s41745-020-00178-1 (2020).

4. Edgar, R., Domrachev, M. & Lash, A. E. Gene Expression Omnibus: NCBI gene expression and hybridization array data repository. Nucleic Acids Res 30, 207–210, https://doi.org/10.1093/nar/30.1.207 (2002).

5. Barrett, T. et al. NCBI GEO: archive for functional genomics data sets–update. Nucleic Acids Res 41, D991–995, https://doi.org/10.1093/nar/gks1193 (2013).

Cited by 15 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Development of a prognostic model for NSCLC based on differential genes in tumour stem cells;Scientific Reports;2024-09-09

2. Prospect of large language models and natural language processing for lung cancer diagnosis: A systematic review;Expert Systems;2024-08-15

3. Single cell transcriptomic analysis reveals tumor immune infiltration by NK cells gene signature in lung adenocarcinoma;Heliyon;2024-07

4. Obesity-associated microbiomes instigate visceral adipose tissue inflammation by recruitment of distinct neutrophils;Nature Communications;2024-06-27

5. Comprehensive single-cell atlas of the mouse retina;iScience;2024-06