Integration of protein sequence and protein–protein interaction data by hypergraph learning to identify novel protein complexes-Reference-Cited by-同舟云学术

Integration of protein sequence and protein–protein interaction data by hypergraph learning to identify novel protein complexes

Published:2024-05-23 Issue:4 Volume:25 Page:
ISSN:1467-5463
Container-title:Briefings in Bioinformatics
language:en
Short-container-title:

Author:

Xia Simin¹²^ORCID,Li Dianke²³⁴^ORCID,Deng Xinru²^ORCID,Liu Zhongyang²^ORCID,Zhu Huaqing¹^ORCID,Liu Yuan²^ORCID,Li Dong¹²^ORCID

Affiliation:

1. School of Basic Medical Sciences, Anhui Medical University , 81 Meishan Road, Shushan District, Hefei 230032 , China

2. State Key Laboratory of Medical Proteomics, Beijing Proteome Research Center, National Center for Protein Sciences (Beijing), Beijing Institute of Lifeomics , 38 Life Science Park, Changping District, Beijing 102206 , China

3. State Key Laboratory of Farm Animal Biotech Breeding , College of Biological Sciences, , 2 Yuanmingyuan West Road, Haidian District, Beijing 100193 , China

4. China Agricultural University , College of Biological Sciences, , 2 Yuanmingyuan West Road, Haidian District, Beijing 100193 , China

Abstract

Abstract Protein–protein interactions (PPIs) are the basis of many important biological processes, with protein complexes being the key forms implementing these interactions. Understanding protein complexes and their functions is critical for elucidating mechanisms of life processes, disease diagnosis and treatment and drug development. However, experimental methods for identifying protein complexes have many limitations. Therefore, it is necessary to use computational methods to predict protein complexes. Protein sequences can indicate the structure and biological functions of proteins, while also determining their binding abilities with other proteins, influencing the formation of protein complexes. Integrating these characteristics to predict protein complexes is very promising, but currently there is no effective framework that can utilize both protein sequence and PPI network topology for complex prediction. To address this challenge, we have developed HyperGraphComplex, a method based on hypergraph variational autoencoder that can capture expressive features from protein sequences without feature engineering, while also considering topological properties in PPI networks, to predict protein complexes. Experiment results demonstrated that HyperGraphComplex achieves satisfactory predictive performance when compared with state-of-art methods. Further bioinformatics analysis shows that the predicted protein complexes have similar attributes to known ones. Moreover, case studies corroborated the remarkable predictive capability of our model in identifying protein complexes, including 3 that were not only experimentally validated by recent studies but also exhibited high-confidence structural predictions from AlphaFold-Multimer. We believe that the HyperGraphComplex algorithm and our provided proteome-wide high-confidence protein complex prediction dataset will help elucidate how proteins regulate cellular processes in the form of complexes, and facilitate disease diagnosis and treatment and drug development. Source codes are available at https://github.com/LiDlab/HyperGraphComplex.

Funder

National Key Research and Development Program of China

National Natural Science Foundation of China

Publisher

Oxford University Press (OUP)

Link

https://academic.oup.com/bib/article-pdf/25/4/bbae274/58173779/bbae274.pdf

Reference72 articles.

1. Recent advances in the development of protein–protein interactions modulators: mechanisms and clinical trials;Lu;Signal Transduct Target Ther,2020

2. Community of protein complexes impacts disease association;Wang;Europ J Hum Genet,2012

3. Protein complexes form a basis for complex hybrid incompatibility;Swamy;Front Genet,2021

4. TOR complexes and the maintenance of cellular homeostasis;Eltschinger;Trends Cell Biol,2016

5. Probing cellular protein complexes using single-molecule pull-down;Jain;Nature,2011