Evolutionary analysis across mammals reveals distinct classes of long non-coding RNAs

被引:125
作者
Chen, Jenny [1 ,2 ]
Shishkin, Alexander A. [3 ]
Zhu, Xiaopeng [4 ]
Kadri, Sabah [1 ]
Maza, Itay [5 ]
Guttman, Mitchell [3 ]
Hanna, Jacob H. [5 ]
Regev, Aviv [1 ,6 ]
Garber, Manuel [4 ,7 ]
机构
[1] Broad Inst MIT & Harvard, Cambridge, MA 02142 USA
[2] MIT, Div Hlth Sci & Technol, Cambridge, MA 02140 USA
[3] CALTECH, Div Biol & Biol Engn, Cambridge, MA 02140 USA
[4] Univ Massachusetts, Program Bioinformat & Integrat Biol, Sch Med, Worcester, MA 01655 USA
[5] Weizmann Inst Sci, Dept Mol Genet, IL-76100 Rehovot, Israel
[6] MIT, Dept Biol, Howard Hughes Med Inst, Cambridge, MA 02140 USA
[7] Univ Massachusetts, Program Mol Biol, Sch Med, Worcester, MA 01655 USA
关键词
Long non-coding RNAs; Evolution; Comparative genomics; Molecular evolution; Annotation; LincRNA; RNA-seq; Transcriptome; GENOME BROWSER DATABASE; EMBRYONIC STEM-CELLS; GENE STRUCTURE; ANNOTATION; DYNAMICS; TRANSCRIPTOMES; PRINCIPLES; ALIGNMENT; LINCRNAS; SEQUENCE;
D O I
10.1186/s13059-016-0880-9
中图分类号
Q81 [生物工程学(生物技术)]; Q93 [微生物学];
学科分类号
071005 ; 0836 ; 090102 ; 100705 ;
摘要
Background: Recent advances in transcriptome sequencing have enabled the discovery of thousands of long non-coding RNAs (lncRNAs) across many species. Though several lncRNAs have been shown to play important roles in diverse biological processes, the functions and mechanisms of most lncRNAs remain unknown. Two significant obstacles lie between transcriptome sequencing and functional characterization of lncRNAs: identifying truly non-coding genes from de novo reconstructed transcriptomes, and prioritizing the hundreds of resulting putative lncRNAs for downstream experimental interrogation. Results: We present slncky, a lncRNA discovery tool that produces a high-quality set of lncRNAs from RNA-sequencing data and further uses evolutionary constraint to prioritize lncRNAs that are likely to be functionally important. Our automated filtering pipeline is comparable to manual curation efforts and more sensitive than previously published computational approaches. Furthermore, we developed a sensitive alignment pipeline for aligning lncRNA loci and propose new evolutionary metrics relevant for analyzing sequence and transcript evolution. Our analysis reveals that evolutionary selection acts in several distinct patterns, and uncovers two notable classes of intergenic lncRNAs: one showing strong purifying selection on RNA sequence and another where constraint is restricted to the regulation but not the sequence of the transcript. Conclusion: Our results highlight that lncRNAs are not a homogenous class of molecules but rather a mixture of multiple functional classes with distinct biological mechanism and/or roles. Our novel comparative methods for lncRNAs reveals 233 constrained lncRNAs out of tens of thousands of currently annotated transcripts, which we make available through the slncky Evolution Browser.
引用
收藏
页数:17
相关论文
共 59 条
[1]  
Bafna V, 2000, Proc Int Conf Intell Syst Mol Biol, V8, P3
[2]   DELETIONS OF THE STEROID SULFATASE GENE IN CLASSICAL X-LINKED ICHTHYOSIS AND IN X-LINKED ICHTHYOSIS ASSOCIATED WITH KALLMANN SYNDROME [J].
BALLABIO, A ;
SEBASTIO, G ;
CARROZZO, R ;
PARENTI, G ;
PICCIRILLO, A ;
PERSICO, MG ;
ANDRIA, G .
HUMAN GENETICS, 1987, 77 (04) :338-341
[3]   Human and mouse gene structure: Comparative analysis and application to exon prediction [J].
Batzoglou, S ;
Pachter, L ;
Mesirov, JP ;
Berger, B ;
Lander, ES .
GENOME RESEARCH, 2000, 10 (07) :950-958
[4]  
Benaglia T, 2009, J STAT SOFTW, V32, P1
[5]   THE PRODUCT OF THE H19 GENE MAY FUNCTION AS AN RNA [J].
BRANNAN, CI ;
DEES, EC ;
INGRAM, RS ;
TILGHMAN, SM .
MOLECULAR AND CELLULAR BIOLOGY, 1990, 10 (01) :28-36
[6]   The evolution of gene expression levels in mammalian organs [J].
Brawand, David ;
Soumillon, Magali ;
Necsulea, Anamaria ;
Julien, Philippe ;
Csardi, Gabor ;
Harrigan, Patrick ;
Weier, Manuela ;
Liechti, Angelica ;
Aximu-Petri, Ayinuer ;
Kircher, Martin ;
Albert, Frank W. ;
Zeller, Ulrich ;
Khaitovich, Philipp ;
Gruetzner, Frank ;
Bergmann, Sven ;
Nielsen, Rasmus ;
Paeaebo, Svante ;
Kaessmann, Henrik .
NATURE, 2011, 478 (7369) :343-+
[7]   Integrative annotation of human large intergenic noncoding RNAs reveals global properties and specific subclasses [J].
Cabili, Moran N. ;
Trapnell, Cole ;
Goff, Loyal ;
Koziol, Magdalena ;
Tazon-Vega, Barbara ;
Regev, Aviv ;
Rinn, John L. .
GENES & DEVELOPMENT, 2011, 25 (18) :1915-1927
[8]   A Long Noncoding RNA Mediates Both Activation and Repression of Immune Response Genes [J].
Carpenter, Susan ;
Aiello, Daniel ;
Atianand, Maninjay K. ;
Ricci, Emiliano P. ;
Gandhi, Pallavi ;
Hall, Lisa L. ;
Byron, Meg ;
Monks, Brian ;
Henry-Bezy, Meabh ;
Lawrence, Jeanne B. ;
O'Neill, Luke A. J. ;
Moore, Melissa J. ;
Caffrey, Daniel R. ;
Fitzgerald, Katherine A. .
SCIENCE, 2013, 341 (6147) :789-792
[9]   The GENCODE v7 catalog of human long noncoding RNAs: Analysis of their gene structure, evolution, and expression [J].
Derrien, Thomas ;
Johnson, Rory ;
Bussotti, Giovanni ;
Tanzer, Andrea ;
Djebali, Sarah ;
Tilgner, Hagen ;
Guernec, Gregory ;
Martin, David ;
Merkel, Angelika ;
Knowles, David G. ;
Lagarde, Julien ;
Veeravalli, Lavanya ;
Ruan, Xiaoan ;
Ruan, Yijun ;
Lassmann, Timo ;
Carninci, Piero ;
Brown, James B. ;
Lipovich, Leonard ;
Gonzalez, Jose M. ;
Thomas, Mark ;
Davis, Carrie A. ;
Shiekhattar, Ramin ;
Gingeras, Thomas R. ;
Hubbard, Tim J. ;
Notredame, Cedric ;
Harrow, Jennifer ;
Guigo, Roderic .
GENOME RESEARCH, 2012, 22 (09) :1775-1789
[10]  
Ellis Blake C., 2012, Frontiers in Genetics, V3, P270, DOI 10.3389/fgene.2012.00270