A data-driven interactome of synergistic genes improves network-based cancer outcome prediction

被引:9
作者
Allahyar, Amin [1 ,2 ]
Ubels, Joske [1 ,3 ,4 ]
de Ridder, Jeroen [1 ]
机构
[1] Univ Utrecht, Univ Med Ctr Utrecht, Dept Genet, Ctr Mol Med, Utrecht, Netherlands
[2] Delft Univ Technol, Delft Bioinformat Lab, Fac Elect Engn Math & Comp Sci, Delft, Netherlands
[3] Skyline DX, Rotterdam, Netherlands
[4] Erasmus MC Canc Inst, Dept Hematol, Rotterdam, Netherlands
关键词
PROTEIN-INTERACTION NETWORKS; BREAST-CANCER; SCALE MAP; EXPRESSION; ARCHITECTURE; SELECTION; DISEASE; KINASE; METASTASIS; VALIDATION;
D O I
10.1371/journal.pcbi.1006657
中图分类号
Q5 [生物化学];
学科分类号
071010 ; 081704 ;
摘要
Robustly predicting outcome for cancer patients from gene expression is an important challenge on the road to better personalized treatment. Network-based outcome predictors (NOPs), which considers the cellular wiring diagram in the classification, hold much promise to improve performance, stability and interpretability of identified marker genes. Problematically, reports on the efficacy of NOPs are conflicting and for instance suggest that utilizing random networks performs on par to networks that describe biologically relevant interactions. In this paper we turn the prediction problem around: instead of using a given biological network in the NOP, we aim to identify the network of genes that truly improves outcome prediction. To this end, we propose SyNet, a gene network constructed ab initio from synergistic gene pairs derived from survival-labelled gene expression data. To obtain SyNet, we evaluate synergy for all 69 million pairwise combinations of genes resulting in a network that is specific to the dataset and phenotype under study and can be used to in a NOP model. We evaluated SyNet and 11 other networks on a compendium dataset of >4000 survival-labelled breast cancer samples. For this purpose, we used cross-study validation which more closely emulates real world application of these outcome predictors. We find that SyNet is the only network that truly improves performance, stability and interpretability in several existing NOPs. We show that SyNet overlaps significantly with existing gene networks, and can be confidently predicted (similar to 85% AUC) from graph-topological descriptions of these networks, in particular the breast tissue-specific network. Due to its data-driven nature, SyNet is not biased to well-studied genes and thus facilitates post-hoc interpretation. We find that SyNet is highly enriched for known breast cancer genes and genes related to e.g. histological grade and tamoxifen resistance, suggestive of a role in determining breast cancer outcome.
引用
收藏
页数:21
相关论文
共 90 条
[51]   Removing Batch Effects from Longitudinal Gene Expression - Quantile Normalization Plus ComBat as Best Approach for Microarray Transcriptome Data [J].
Mueller, Christian ;
Schillert, Arne ;
Roethemeier, Caroline ;
Tregouet, David-Alexandre ;
Proust, Carole ;
Binder, Harald ;
Pfeiffer, Norbert ;
Beutel, Manfred ;
Lackner, Karl J. ;
Schnabel, Renate B. ;
Tiret, Laurence ;
Wild, Philipp S. ;
Blankenberg, Stefan ;
Zeller, Tanja ;
Ziegler, Andreas .
PLOS ONE, 2016, 11 (06)
[52]   Maternal embryonic leucine zipper kinase (MELK) regulates multipotent neural progenitor proliferation [J].
Nakano, I ;
Paucar, AA ;
Bajpai, R ;
Dougherty, JD ;
Zewail, A ;
Kelly, TK ;
Kim, KJ ;
Ou, J ;
Groszer, M ;
Imura, T ;
Freije, WA ;
Nelson, SF ;
Sofroniew, MV ;
Wu, H ;
Liu, X ;
Terskikh, AV ;
Geschwind, DH ;
Kornblum, HI .
JOURNAL OF CELL BIOLOGY, 2005, 170 (03) :413-427
[53]   The tumor suppressor CDKN3 controls mitosis [J].
Nalepa, Grzegorz ;
Barnholtz-Sloan, Jill ;
Enzor, Rikki ;
Dey, Dilip ;
He, Ying ;
Gehlhausen, Jeff R. ;
Lehmann, Amalia S. ;
Park, Su-Jung ;
Yang, Yanzhu ;
Yang, Xianlin ;
Chen, Shi ;
Guan, Xiaowei ;
Chen, Yanwen ;
Renbarger, Jamie ;
Yang, Feng-Chun ;
Parada, Luis F. ;
Clapp, Wade .
JOURNAL OF CELL BIOLOGY, 2013, 201 (07) :997-1012
[54]  
Newman M, 2004, PHYS REV E, V69, P1, DOI DOI 10.1103/PhysRevE.69.026113
[55]   Averaged gene expressions for regression [J].
Park, Mee Young ;
Hastie, Trevor ;
Tibshirani, Robert .
BIOSTATISTICS, 2007, 8 (02) :212-227
[56]  
Parker Hilary S., 2012, STAT APPL GENETICS M
[57]   Cooperative assembly of CYK-4/MgcRacGAP and ZEN-4/MKLP1 to form the centralspindlin complex [J].
Pavicic-Kaltenbrunner, Visnja ;
Mishima, Masanori ;
Glotzer, Michael .
MOLECULAR BIOLOGY OF THE CELL, 2007, 18 (12) :4992-5003
[58]   Effect of training-sample size and classification difficulty on the accuracy of genomic predictors [J].
Popovici, Vlad ;
Chen, Weijie ;
Gallas, Brandon G. ;
Hatzis, Christos ;
Shi, Weiwei ;
Samuelson, Frank W. ;
Nikolsky, Yuri ;
Tsyganova, Marina ;
Ishkin, Alex ;
Nikolskaya, Tatiana ;
Hess, Kenneth R. ;
Valero, Vicente ;
Booser, Daniel ;
Delorenzi, Mauro ;
Hortobagyi, Gabriel N. ;
Shi, Leming ;
Symmans, W. Fraser ;
Pusztai, Lajos .
BREAST CANCER RESEARCH, 2010, 12 (01)
[59]   Human Protein Reference Database-2009 update [J].
Prasad, T. S. Keshava ;
Goel, Renu ;
Kandasamy, Kumaran ;
Keerthikumar, Shivakumar ;
Kumar, Sameer ;
Mathivanan, Suresh ;
Telikicherla, Deepthi ;
Raju, Rajesh ;
Shafreen, Beema ;
Venugopal, Abhilash ;
Balakrishnan, Lavanya ;
Marimuthu, Arivusudar ;
Banerjee, Sutopa ;
Somanathan, Devi S. ;
Sebastian, Aimy ;
Rani, Sandhya ;
Ray, Somak ;
Kishore, C. J. Harrys ;
Kanth, Sashi ;
Ahmed, Mukhtar ;
Kashyap, Manoj K. ;
Mohmood, Riaz ;
Ramachandra, Y. L. ;
Krishna, V. ;
Rahiman, B. Abdul ;
Mohan, Sujatha ;
Ranganathan, Prathibha ;
Ramabadran, Subhashri ;
Chaerkady, Raghothama ;
Pandey, Akhilesh .
NUCLEIC ACIDS RESEARCH, 2009, 37 :D767-D772
[60]   Breast cancer prognostic classification in the molecular era: the role of histological grade [J].
Rakha, Emad A. ;
Reis-Filho, Jorge S. ;
Baehner, Frederick ;
Dabbs, David J. ;
Decker, Thomas ;
Eusebi, Vincenzo ;
Fox, Stephen B. ;
Ichihara, Shu ;
Jacquemier, Jocelyne ;
Lakhani, Sunil R. ;
Palacios, Jose ;
Richardson, Andrea L. ;
Schnitt, Stuart J. ;
Schmitt, Fernando C. ;
Tan, Puay-Hoon ;
Tse, Gary M. ;
Badve, Sunil ;
Ellis, Ian O. .
BREAST CANCER RESEARCH, 2010, 12 (04)