Node-adaptive graph Transformer with structural encoding for accurate and robust lncRNA-disease association prediction

Author:

Li Guanghui1,Bai Peihao1,Liang Cheng2,Luo Jiawei3

Affiliation:

1. East China Jiaotong University

2. Shandong Normal University

3. Hunan University

Abstract

Abstract Background Long noncoding RNAs (lncRNAs) are integral to a plethora of critical cellular biological processes, including the regulation of gene expression, cell differentiation, and the development of tumors and cancers. Predicting the relationships between lncRNAs and diseases can contribute to a better understanding of the pathogenic mechanisms of disease and provide strong support for the development of advanced treatment methods.Results Therefore, we present an innovative node-adaptive Transformer model for predicting unknown associations between lncRNAs and diseases (GNATLDA). First, we utilize the node-adaptive feature smoothing (NAFS) method to learn the local feature information of nodes and encode the structural information of the fusion similarity network of diseases and lncRNAs using Structural Deep Network Embedding (SDNE). Next, the Transformer module, which contains a multi-headed attention layer, is used to learn global feature information about the nodes of the heterogeneous network, which is used to capture potential association information between the network nodes. Finally, we employ a Transformer module with two multi-headed attention layers for learning global-level embedding fusion. Network structure coding is added as the structural inductive bias of the network to compensate for the missing message-passing mechanism in Transformer. Our model accounts for both local-level and global-level node information and exploits the global horizon of the Transformer model, which fuses the structural inductive bias of the network to comprehensively investigate unidentified associations between nodes, significantly increasing the predictive effectiveness of potential interactions between diseases and lncRNAs. We conducted case studies on four diseases; 55 out of 60 interactions between diseases and lncRNAs were confirmed by the literature.Conclusions Our proposed GNATLDA model can serve as a highly efficient computational method for predicting biological information associations.

Publisher

Research Square Platform LLC

Reference68 articles.

1. The GENCODE v7 catalog of human long noncoding RNAs: Analysis of their gene structure, evolution, and expression;Derrien T;Genome Res,2012

2. Modular regulatory principles of large non-coding RNAs;Guttman M;Nature,2012

3. Molecular Mechanisms of Long Noncoding RNAs;Wang Kevin C;Mol Cell,2011

4. Long noncoding RNAs and human disease;Wapinski O;Trends Cell Biol,2011

5. Long non-coding RNAs and complex diseases: from experimental results to computational models;Chen X;Brief Bioinform,2016

同舟云学术

1.学者识别学者识别

2.学术分析学术分析

3.人才评估人才评估

"同舟云学术"是以全球学者为主线,采集、加工和组织学术论文而形成的新型学术文献查询和分析系统,可以对全球学者进行文献检索和人才价值评估。用户可以通过关注某些学科领域的顶尖人物而持续追踪该领域的学科进展和研究前沿。经过近期的数据扩容,当前同舟云学术共收录了国内外主流学术期刊6万余种,收集的期刊论文及会议论文总量共计约1.5亿篇,并以每天添加12000余篇中外论文的速度递增。我们也可以为用户提供个性化、定制化的学者数据。欢迎来电咨询!咨询电话:010-8811{复制后删除}0370

www.globalauthorid.com

TOP

Copyright © 2019-2024 北京同舟云网络信息技术有限公司
京公网安备11010802033243号  京ICP备18003416号-3