Tensor-Decomposition-Based Unsupervised Feature Extraction Applied to Prostate Cancer Multiomics Data-Reference-Cited by-同舟云学术

Tensor-Decomposition-Based Unsupervised Feature Extraction Applied to Prostate Cancer Multiomics Data

Published:2020-12-11 Issue:12 Volume:11 Page:1493
ISSN:2073-4425
Container-title:Genes
language:en
Short-container-title:Genes

Author:

Taguchi Y-h.^ORCID,Turki Turki^ORCID

Abstract

The large p small n problem is a challenge without a de facto standard method available to it. In this study, we propose a tensor-decomposition (TD)-based unsupervised feature extraction (FE) formalism applied to multiomics datasets, in which the number of features is more than 100,000 whereas the number of samples is as small as about 100, hence constituting a typical large p small n problem. The proposed TD-based unsupervised FE outperformed other conventional supervised feature selection methods, random forest, categorical regression (also known as analysis of variance, or ANOVA), penalized linear discriminant analysis, and two unsupervised methods, multiple non-negative matrix factorization and principal component analysis (PCA) based unsupervised FE when applied to synthetic datasets and four methods other than PCA based unsupervised FE when applied to multiomics datasets. The genes selected by TD-based unsupervised FE were enriched in genes known to be related to tissues and transcription factors measured. TD-based unsupervised FE was demonstrated to be not only the superior feature selection method but also the method that can select biologically reliable genes. To our knowledge, this is the first study in which TD-based unsupervised FE has been successfully applied to the integration of this variety of multiomics measurements.

Publisher

MDPI AG

Subject

Genetics (clinical),Genetics

Link

https://www.mdpi.com/2073-4425/11/12/1493/pdf

Reference51 articles.

1. Efficient learning from big data for cancer risk modeling: A case study with melanoma

2. GPU-DAEMON: GPU algorithm design, data management & optimization template for array based big omics data;Awan;Comput. Biol. Med.,2018

3. Scaling up Machine Learning: Parallel and Distributed Approaches;Bekkerman,2011

Cited by 4 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Application of TD-Based Unsupervised FE to Bioinformatics;Unsupervised and Semi-Supervised Learning;2024

2. From molecular mechanisms of prostate cancer to translational applications: based on multi-omics fusion analysis and intelligent medicine;Health Information Science and Systems;2023-12-18

3. Unsupervised tensor decomposition-based method to extract candidate transcription factors as histone modification bookmarks in post-mitotic transcriptional reactivation;PLOS ONE;2021-05-25

4. Unsupervised tensor decomposition-based method to extract candidate transcription factors as histone modification bookmarks in post-mitotic transcriptional reactivation;2020-09-24