APIS: accurate prediction of hot spots in protein interfaces by combining protrusion index with solvent accessibility-Reference-Cited by-同舟云学术

APIS: accurate prediction of hot spots in protein interfaces by combining protrusion index with solvent accessibility

Published:2010-04-08 Issue:1 Volume:11 Page:
ISSN:1471-2105
Container-title:BMC Bioinformatics
language:en
Short-container-title:BMC Bioinformatics

Author:

Xia Jun-Feng,Zhao Xing-Ming,Song Jiangning,Huang De-Shuang

Abstract

Abstract Background It is well known that most of the binding free energy of protein interaction is contributed by a few key hot spot residues. These residues are crucial for understanding the function of proteins and studying their interactions. Experimental hot spots detection methods such as alanine scanning mutagenesis are not applicable on a large scale since they are time consuming and expensive. Therefore, reliable and efficient computational methods for identifying hot spots are greatly desired and urgently required. Results In this work, we introduce an efficient approach that uses support vector machine (SVM) to predict hot spot residues in protein interfaces. We systematically investigate a wide variety of 62 features from a combination of protein sequence and structure information. Then, to remove redundant and irrelevant features and improve the prediction performance, feature selection is employed using the F-score method. Based on the selected features, nine individual-feature based predictors are developed to identify hot spots using SVMs. Furthermore, a new ensemble classifier, namely APIS (A combined model based on Protrusion Index and Solvent accessibility), is developed to further improve the prediction accuracy. The results on two benchmark datasets, ASEdb and BID, show that this proposed method yields significantly better prediction accuracy than those previously published in the literature. In addition, we also demonstrate the predictive power of our proposed method by modelling two protein complexes: the calmodulin/myosin light chain kinase complex and the heat shock locus gene products U and V complex, which indicate that our method can identify more hot spots in these two complexes compared with other state-of-the-art methods. Conclusion We have developed an accurate prediction model for hot spot residues, given the structure of a protein complex. A major contribution of this study is to propose several new features based on the protrusion index of amino acid residues, which has been shown to significantly improve the prediction performance of hot spots. Moreover, we identify a compact and useful feature subset that has an important implication for identifying hot spot residues. Our results indicate that these features are more effective than the conventional evolutionary conservation, pairwise residue potentials and other traditional features considered previously, and that the combination of our and traditional features may support the creation of a discriminative feature set for efficient prediction of hot spot residues. The data and source code are available on web site http://home.ustc.edu.cn/~jfxia/hotspot.html.

Publisher

Springer Science and Business Media LLC

Subject

Applied Mathematics,Computer Science Applications,Molecular Biology,Biochemistry,Structural Biology

Link

https://link.springer.com/content/pdf/10.1186/1471-2105-11-174.pdf

Reference57 articles.

1. Wu Z, Zhao X, Chen L: Identifying responsive functional modules from protein-protein interaction network. Molecules and Cells 2009, 27(3):271–277. 10.1007/s10059-009-0035-x

2. Zhao X, Wang R, Chen L, Aihara K: Uncovering signal transduction networks from high-throughput data by integer linear programming. Nucleic Acids Research 2008, 36(9):e48. 10.1093/nar/gkn145

3. Xia J, Han K, Huang D: Sequence-Based Prediction of Protein-Protein Interactions by Means of Rotation Forest and Autocorrelation Descriptor. Protein and Peptide Letters 2010, 17(1):137–145. 10.2174/092986610789909403

4. Zhao X, Chen L, Aihara K: A discriminative approach to identifying domain-domain interactions from protein-protein interactions. Proteins 2010, 78(5):1243–1253. 10.1002/prot.22643

5. Moreira I, Fernandes P, Ramos M: Hot spots--A review of the protein-protein interface determinant amino-acid residues. Proteins 2007, 68: 803–812. 10.1002/prot.21396

Cited by 188 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. PPI-hotspotID: A Method for Detecting Protein-Protein Interaction Hot Spots from the Free Protein Structure;2024-08-07

2. PPI-hotspotID: A Method for Detecting Protein-Protein Interaction Hot Spots from the Free Protein Structure;2024-08-07

3. PPI-hotspotID: A Method for Detecting Protein-Protein Interaction Hot Spots from the Free Protein Structure;2024-05-28

4. PPI-hotspotID: A Method for Detecting Protein-Protein Interaction Hot Spots from the Free Protein Structure;2024-05-28

5. Prediction of Protein-DNA Interface Hot Spots Based on Empirical Mode Decomposition and Machine Learning;Genes;2024-05-23