Machine Learning and Integrative Analysis of Biomedical Big Data-Reference-Cited by-同舟云学术

Machine Learning and Integrative Analysis of Biomedical Big Data

Published:2019-01-28 Issue:2 Volume:10 Page:87
ISSN:2073-4425
Container-title:Genes
language:en
Short-container-title:Genes

Author:

Mirza Bilal,Wang Wei,Wang Jie,Choi Howard,Chung Neo Christopher^ORCID,Ping Peipei

Abstract

Recent developments in high-throughput technologies have accelerated the accumulation of massive amounts of omics data from multiple sources: genome, epigenome, transcriptome, proteome, metabolome, etc. Traditionally, data from each source (e.g., genome) is analyzed in isolation using statistical and machine learning (ML) methods. Integrative analysis of multi-omics and clinical data is key to new biomedical discoveries and advancements in precision medicine. However, data integration poses new computational challenges as well as exacerbates the ones associated with single-omics studies. Specialized computational approaches are required to effectively and efficiently perform integrative analysis of biomedical data acquired from diverse modalities. In this review, we discuss state-of-the-art ML-based approaches for tackling five specific computational challenges associated with integrative analysis: curse of dimensionality, data heterogeneity, missing data, class imbalance and scalability issues.

Funder

National Institutes of Health

Publisher

MDPI AG

Subject

Genetics (clinical),Genetics

Link

http://www.mdpi.com/2073-4425/10/2/87/pdf

Reference233 articles.

1. High-throughput determination of RNA structures

2. Single-cell RNA sequencing technologies and bioinformatics pipelines

3. Piercing the dark matter: bioinformatics of long-range sequencing and mapping