Learning the Regulatory Code of Gene Expression-Reference-Cited by-同舟云学术

Learning the Regulatory Code of Gene Expression

Published:2021-06-10 Issue: Volume:8 Page:
ISSN:2296-889X
Container-title:Frontiers in Molecular Biosciences
language:
Short-container-title:Front. Mol. Biosci.

Author:

Zrimec Jan,Buric Filip,Kokina Mariia,Garcia Victor,Zelezniak Aleksej

Abstract

Data-driven machine learning is the method of choice for predicting molecular phenotypes from nucleotide sequence, modeling gene expression events including protein-DNA binding, chromatin states as well as mRNA and protein levels. Deep neural networks automatically learn informative sequence representations and interpreting them enables us to improve our understanding of the regulatory code governing gene expression. Here, we review the latest developments that apply shallow or deep learning to quantify molecular phenotypes and decode the cis-regulatory grammar from prokaryotic and eukaryotic sequencing data. Our approach is to build from the ground up, first focusing on the initiating protein-DNA interactions, then specific coding and non-coding regions, and finally on advances that combine multiple parts of the gene and mRNA regulatory structures, achieving unprecedented performance. We thus provide a quantitative view of gene expression regulation from nucleotide sequence, concluding with an information-centric overview of the central dogma of molecular biology.

Funder

Vetenskapsrådet

Publisher

Frontiers Media SA

Subject

Biochemistry, Genetics and Molecular Biology (miscellaneous),Molecular Biology,Biochemistry

Reference287 articles.

1. Deconvolving the Recognition of DNA Shape from Sequence;Abe;Cell,2015

2. Predicting mRNA Abundance Directly from Genomic Sequence Using Deep Convolutional Neural Networks;Agarwal;Cell Rep,2020

3. Predicting the Sequence Specificities of DNA- and RNA-Binding Proteins by Deep Learning;Alipanahi;Nat. Biotechnol.,2015

4. DeepCpG: Accurate Prediction of Single-Cell DNA Methylation States Using Deep Learning;Angermueller;Genome Biol.,2017

Cited by 22 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Deep Generative Optimization of mRNA Codon Sequences for Enhanced Protein Production and Therapeutic Efficacy;2024-09-08

2. CBLANE: A deep learning approach for Transcription Factor Binding Sites Prediction;2024-05-22

3. Comparative Analysis of DNA Structural Parameters and the Corresponding Computational Tools to Differentiate Regulatory DNA Motifs and Promoters;2024-03-27

4. Artificial Intelligence and Machine Learning in Bioinformatics;Reference Module in Life Sciences;2024

5. Promoters in Pichia pastoris: A Toolbox for Fine-Tuned Gene Expression;Methods in Molecular Biology;2024