Computational approaches for detecting protein complexes from protein interaction networks: a survey

被引:227
作者
Li, Xiaoli [1 ]
Wu, Min [2 ]
Kwoh, Chee-Keong [2 ]
Ng, See-Kiong [1 ]
机构
[1] Inst Infocomm Res, Singapore, Singapore
[2] Nanyang Technol Univ, Sch Comp Engn, Singapore, Singapore
关键词
FUNCTIONAL MODULES; SEMANTIC SIMILARITY; PHYSICAL INTERACTOME; YEAST; ALGORITHM; INTEGRATION; ANNOTATION; PREDICTION; MODULARITY; TOPOLOGY;
D O I
10.1186/1471-2164-11-S1-S3
中图分类号
Q81 [生物工程学(生物技术)]; Q93 [微生物学];
学科分类号
071005 ; 0836 ; 090102 ; 100705 ;
摘要
Background: Most proteins form macromolecular complexes to perform their biological functions. However, experimentally determined protein complex data, especially of those involving more than two protein partners, are relatively limited in the current state-of-the-art high-throughput experimental techniques. Nevertheless, many techniques (such as yeast-two-hybrid) have enabled systematic screening of pairwise protein-protein interactions en masse. Thus computational approaches for detecting protein complexes from protein interaction data are useful complements to the limited experimental methods. They can be used together with the experimental methods for mapping the interactions of proteins to understand how different proteins are organized into higher-level substructures to perform various cellular functions. Results: Given the abundance of pairwise protein interaction data from high-throughput genome-wide experimental screenings, a protein interaction network can be constructed from protein interaction data by considering individual proteins as the nodes, and the existence of a physical interaction between a pair of proteins as a link. This binary protein interaction graph can then be used for detecting protein complexes using graph clustering techniques. In this paper, we review and evaluate the state-of-the-art techniques for computational detection of protein complexes, and discuss some promising research directions in this field. Conclusions: Experimental results with yeast protein interaction data show that the interaction subgraphs discovered by various computational methods matched well with actual protein complexes. In addition, the computational approaches have also improved in performance over the years. Further improvements could be achieved if the quality of the underlying protein interaction data can be considered adequately to minimize the undesirable effects from the irrelevant and noisy sources, and the various biological evidences can be better incorporated into the detection process to maximize the exploitation of the increasing wealth of biological knowledge available.
引用
收藏
页数:19
相关论文
共 106 条
[31]   The ins and outs of signalling [J].
Downward, J .
NATURE, 2001, 411 (6839) :759-762
[32]  
DUTKOWSKI J, 2007, ISMB S, V23, P149
[33]   Saccharomyces Genome Database (SGD) provides secondary gene annotation using the Gene Ontology (GO) [J].
Dwight, SS ;
Harris, MA ;
Dolinski, K ;
Ball, CA ;
Binkley, G ;
Christie, KR ;
Fisk, DG ;
Issel-Tarver, L ;
Schroeder, M ;
Sherlock, G ;
Sethuraman, A ;
Weng, S ;
Botstein, D ;
Cherry, JM .
NUCLEIC ACIDS RESEARCH, 2002, 30 (01) :69-72
[34]   Cluster analysis and display of genome-wide expression patterns [J].
Eisen, MB ;
Spellman, PT ;
Brown, PO ;
Botstein, D .
PROCEEDINGS OF THE NATIONAL ACADEMY OF SCIENCES OF THE UNITED STATES OF AMERICA, 1998, 95 (25) :14863-14868
[35]  
Ester M., 1996, DENSITY BASED ALGORI, DOI DOI 10.5555/3001460.3001507
[36]   The nuclear pore complex: Nucleocytoplasmic transport and beyond [J].
Fahrenkrog, B ;
Aebi, U .
NATURE REVIEWS MOLECULAR CELL BIOLOGY, 2003, 4 (10) :757-766
[37]  
Feng Jianxing, 2008, Comput Syst Bioinformatics Conf, V7, P51, DOI 10.1142/9781848162648_0005
[38]  
Friedel CC, 2008, LECT N BIOINFORMAT, V4955, P3
[39]   A FAST PARAMETRIC MAXIMUM FLOW ALGORITHM AND APPLICATIONS [J].
GALLO, G ;
GRIGORIADIS, MD ;
TARJAN, RE .
SIAM JOURNAL ON COMPUTING, 1989, 18 (01) :30-55
[40]  
Garey M.R., 1979, SER MATH SCI SERIES