Inference of differential gene regulatory networks from gene expression data using boosted differential trees

Author:

Galindez GihannaORCID,List MarkusORCID,Baumbach Jan,Blumenthal David B.ORCID,Kacprowski Tim

Abstract

AbstractDiseases can be caused by molecular perturbations that induce specific changes in regulatory interactions and their coordinated expression, also referred to as network rewiring. However, the detection of complex changes in regulatory connections remains a challenging task and would benefit from the development of novel non-parametric approaches. We developed a new ensemble method called BoostDiff (boosted differential regression trees) to infer a differential network discriminating between two conditions. BoostDiff builds an adaptively boosted (AdaBoost) ensemble of differential trees with respect to a target condition. To build the differential trees, we propose differential variance improvement as a novel splitting criterion. Variable importance measures derived from the resulting models are used to reflect changes in gene expression predictability and to build the output differential networks. BoostDiff outperforms existing differential network methods on simulated data evaluated in two different complexity settings. We then demonstrate the power of our approach when applied to real transcriptomics data in COVID-19 and Crohn’s disease. BoostDiff identifies context-specific networks that are enriched with genes of known disease-relevant pathways and complements standard differential expression analyses. BoostDiff is available at https://github.com/gihannagalindez/boostdiff_inference.Author SummaryGene regulatory networks, which comprise the collection of regulatory relationships between transcription factors and their target genes, are important for controlling various molecular processes. Diseases can induce perturbations in normal gene co-expression patterns in these networks. Detecting differentially co-expressed or rewired edges between disease and healthy biological states can be thus useful for investigating the link between specific disease-associated molecular alterations and phenotype. We developed BoostDiff (boosted differential trees), an ensemble method to derive differential networks between two biological contexts. Our approach applies a boosting scheme using differential trees as base learner. A differential tree is a new tree structure that is built from two expression datasets using a splitting criterion called the differential variance improvement. The resulting BoostDiff model learns the most differentially predictive features which are then used to build the directed differential networks. BoostDiff outperforms other differential network methods on simulated data and outputs more biologically meaningful results when evaluated on real transcriptomics datasets. BoostDiff can be applied to gene expression data to reveal new disease mechanisms or identify potential therapeutic targets.

Publisher

Cold Spring Harbor Laboratory

同舟云学术

1.学者识别学者识别

2.学术分析学术分析

3.人才评估人才评估

"同舟云学术"是以全球学者为主线,采集、加工和组织学术论文而形成的新型学术文献查询和分析系统,可以对全球学者进行文献检索和人才价值评估。用户可以通过关注某些学科领域的顶尖人物而持续追踪该领域的学科进展和研究前沿。经过近期的数据扩容,当前同舟云学术共收录了国内外主流学术期刊6万余种,收集的期刊论文及会议论文总量共计约1.5亿篇,并以每天添加12000余篇中外论文的速度递增。我们也可以为用户提供个性化、定制化的学者数据。欢迎来电咨询!咨询电话:010-8811{复制后删除}0370

www.globalauthorid.com

TOP

Copyright © 2019-2024 北京同舟云网络信息技术有限公司
京公网安备11010802033243号  京ICP备18003416号-3