iHSP-PseRAAAC: Identifying the heat shock protein families using pseudo reduced amino acid alphabet composition

被引:279
作者
Feng, Peng-Mian [1 ]
Chen, Wei [2 ,3 ,4 ]
Lin, Hao [5 ]
Chou, Kuo-Chen [4 ,6 ]
机构
[1] Hebei United Univ, Sch Publ Hlth, Tangshan 063000, Peoples R China
[2] Hebei United Univ, Sch Sci, Dept Phys, Tangshan 063000, Peoples R China
[3] Hebei United Univ, Ctr Genom & Computat Biol, Tangshan 063000, Peoples R China
[4] Gordon Life Sci Inst, Belmont, MA 02478 USA
[5] Univ Elect Sci & Technol China, Sch Life Sci & Technol, Ctr Bioinformat, Key Lab Neuroinformat,Minist Educ, Chengdu 610054, Peoples R China
[6] King Abdulaziz Univ, CEGMR, Jeddah 21413, Saudi Arabia
关键词
Heat shock protein; Reduced amino acid alphabet; n-Peptide composition; PseAAC; SVM; Web server; SUPPORT VECTOR MACHINES; PREDICTING SUBCELLULAR-LOCALIZATION; GENERAL-FORM; PHYSICOCHEMICAL PROPERTIES; STRUCTURAL CLASS; SIGNAL PEPTIDES; CHOUS PSEAAC; HEAT-SHOCK-PROTEIN-70; ATTRIBUTES; MODE;
D O I
10.1016/j.ab.2013.05.024
中图分类号
Q5 [生物化学];
学科分类号
071010 ; 081704 ;
摘要
Heat shock proteins (HSPs) are a type of functionally related proteins present in all living organisms, both prokaryotes and eukaryotes. They play essential roles in protein-protein interactions such as folding and assisting in the establishment of proper protein conformation and prevention of unwanted protein aggregation. Their dysfunction may cause various life-threatening disorders, such as Parkinson's, Alzheimer's, and cardiovascular diseases. Based on their functions, HSPs are usually classified into six families: (i) HSP20 or sHSP, (ii) HSP40 or J-class proteins, (iii) HSP60 or GroEL/ES, (iv) HSP70, (v) HSP90, and (vi) HSP100. Although considerable progress has been achieved in discriminating HSPs from other proteins, it is still a big challenge to identify HSPs among their six different functional types according to their sequence information alone. With the avalanche of protein sequences generated in the post-genomic age, it is highly desirable to develop a high-throughput computational tool in this regard. To take up such a challenge, a predictor called iHSP-PseRAAAC has been developed by incorporating the reduced amino acid alphabet information into the general form of pseudo amino acid composition. One of the remarkable advantages of introducing the reduced amino acid alphabet is being able to avoid the notorious dimension disaster or overfitting problem in statistical prediction. It was observed that the overall success rate achieved by iHSP-PseRAAAC in identifying the functional types of HSPs among the aforementioned six types was more than 87%, which was derived by the jackknife test on a stringent benchmark dataset in which none of HSPs included has >= 40% pairwise sequence identity to any other in the same subset. It has not escaped our notice that the reduced amino acid alphabet approach can also be used to investigate other protein classification problems. As a user-friendly web server, iHSP-PseRAAAC is accessible to the public at http://lin.uestc.edu.cn/server/iHSP-PseRAAAC. (C) 2013 Elsevier Inc. All rights reserved.
引用
收藏
页码:118 / 125
页数:8
相关论文
共 86 条
[61]   Universally conserved positions in protein folds: Reading evolutionary signals about stability, folding kinetics and function [J].
Mirny, LA ;
Shakhnovich, EI .
JOURNAL OF MOLECULAR BIOLOGY, 1999, 291 (01) :177-196
[62]   Prediction of Allergenic Proteins by Means of the Concept of Chou's Pseudo Amino Acid Composition and a Machine Learning Approach [J].
Mohabatkar, Hassan ;
Beigi, Majid Mohammad ;
Abdolahi, Kolsoum ;
Mohsenzadeh, Sasan .
MEDICINAL CHEMISTRY, 2013, 9 (01) :133-137
[63]   Prediction of GABAA receptor proteins using the concept of Chou's pseudo-amino acid composition and support vector machine [J].
Mohabatkar, Hassan ;
Beigi, Majid Mohammad ;
Esmaeili, Abolghasem .
JOURNAL OF THEORETICAL BIOLOGY, 2011, 281 (01) :18-23
[64]   Prediction of Cyclin Proteins Using Chou's Pseudo Amino Acid Composition [J].
Mohabatkar, Hassan .
PROTEIN AND PEPTIDE LETTERS, 2010, 17 (10) :1207-1214
[65]  
Mohammad Beigi Majid, 2011, J Struct Funct Genomics, V12, P191, DOI 10.1007/s10969-011-9120-4
[66]   THE FOLDING TYPE OF A PROTEIN IS RELEVANT TO THE AMINO-ACID-COMPOSITION [J].
NAKASHIMA, H ;
NISHIKAWA, K ;
OOI, T .
JOURNAL OF BIOCHEMISTRY, 1986, 99 (01) :153-162
[67]   Genetic programming for creating Chou's pseudo amino acid based features for submitochondria localization [J].
Nanni, Loris ;
Lumini, Alessandra .
AMINO ACIDS, 2008, 34 (04) :653-660
[68]   Identifying Bacterial Virulent Proteins by Fusing a Set of Classifiers Based on Variants of Chou's Pseudo Amino Acid Composition and on Evolutionary Information [J].
Nanni, Loris ;
Lumini, Alessandra ;
Gupta, Dinesh ;
Garg, Aarti .
IEEE-ACM TRANSACTIONS ON COMPUTATIONAL BIOLOGY AND BIOINFORMATICS, 2012, 9 (02) :467-475
[69]   Heat shock proteins, inflammation, and cardiovascular disease [J].
Pockley, AG .
CIRCULATION, 2002, 105 (08) :1012-1017
[70]  
RITOSSA P., 1962, RIV 1ST SIEROTERAPITAL, V37, P79