Abstract
AbstractHigh dimensional genomic data can be analyzed to understand the effects of multiple variables on a target variable such as a clinical outcome, risk factor or diagnosis. Of special interest are functional modules, cooperating sets of variables affecting the target. Graphical models of various types are often useful for characterizing such networks of variables. In other applications such as social networks, the concept of balance in undirected signed graphs characterizes the consistency of associations within the network. To extend this concept to applications where a set of predictor variables influences an outcome variable, we define balance for functional modules. This property specifies that the module variables have a joint effect on the target outcome with no internal conflict, an efficiency that evolution may use for selection in biological networks. We show that for this class of graphs, observed correlations directly reflect paths in the underlying graph. Consequences of the balance property are exploited to implement a new module discovery algorithm, bFMD, which selects a subset of variables from highdimensional data that compose a balanced functional module. Our bFMD algorithm performed favorably in simulations as compared to other module detection methods that do not consider balance properties. Additionally, bFMD detected interpretable results in a real application for RNA-seq data obtained from The Cancer Genome Atlas (TCGA) for Uterine Corpus Endometrial Carcinoma using the percentage of tumor invasion as the target outcome of interest. bFMD detects sparse sets of variables within highdimensional datasets such that interpretability may be favorable as compared to other similar methods by leveraging balance properties used in other graphical applications.
Publisher
Cold Spring Harbor Laboratory
Reference40 articles.
1. Connecting the dots:econometric methods for uncoverring networks with an application to the australian financial institutions;Journal of Banking and Finance,2015
2. Prediction by Supervised Principal Components
3. Semi-Supervised Methods to Predict Patient Survival from Gene Expression Data
4. Surgical Staging in Endometrial Cancer: Clinical-pathologic Findings of a Prospective Study;Obstetrics and Gynecology,1984
5. Independent filtering increases detection power for high-throughput experiments
Cited by
2 articles.
订阅此论文施引文献
订阅此论文施引文献,注册后可以免费订阅5篇论文的施引文献,订阅后可以查看论文全部施引文献