Araport11: a complete reannotation of the Arabidopsis thaliana reference genome

被引:489
作者
Cheng, Chia-Yi [1 ]
Krishnakumar, Vivek [1 ]
Chan, Agnes P. [1 ]
Thibaud-Nissen, Francoise [2 ]
Schobel, Seth [1 ]
Town, Christopher D. [1 ]
机构
[1] J Craig Venter Inst, 9714 Med Ctr Dr, Rockville, MD 20850 USA
[2] NIH, Natl Ctr Biotechnol Informat, US Natl Lib Med, Bethesda, MD 20894 USA
基金
美国国家科学基金会;
关键词
Arabidopsis; annotation; transcriptome; NATURAL ANTISENSE TRANSCRIPTS; NONSENSE-MEDIATED DECAY; SMALL RNA LOCI; GENE-EXPRESSION; WIDE ANALYSIS; COMPREHENSIVE ANNOTATION; NONCODING RNAS; POLYMERASE-IV; SEQ DATA; REVEALS;
D O I
10.1111/tpj.13415
中图分类号
Q94 [植物学];
学科分类号
071001 ;
摘要
The flowering plant Arabidopsis thaliana is a dicot model organism for research in many aspects of plant biology. A comprehensive annotation of its genome paves the way for understanding the functions and activities of all types of transcripts, including mRNA, the various classes of non-coding RNA, and small RNA. The TAIR10 annotation update had a profound impact on Arabidopsis research but was released more than 5years ago. Maintaining the accuracy of the annotation continues to be a prerequisite for future progress. Using an integrative annotation pipeline, we assembled tissue-specific RNA-Seq libraries from 113 datasets and constructed 48359 transcript models of protein-coding genes in eleven tissues. In addition, we annotated various classes of non-coding RNA including microRNA, long intergenic RNA, small nucleolar RNA, natural antisense transcript, small nuclear RNA, and small RNA using published datasets and in-house analytic results. Altogether, we identified 635 novel protein-coding genes, 508 novel transcribed regions, 5178 non-coding RNAs, and 35846 small RNA loci that were formerly unannotated. Analysis of the splicing events and RNA-Seq based expression profiles revealed the landscapes of gene structures, untranslated regions, and splicing activities to be more intricate than previously appreciated. Furthermore, we present 692 uniformly expressed housekeeping genes, 43% of whose human orthologs are also housekeeping genes. This updated Arabidopsis genome annotation with a substantially increased resolution of gene models will not only further our understanding of the biological processes of this plant model but also of other species.
引用
收藏
页码:789 / 804
页数:16
相关论文
共 105 条
[11]   Everything old is new again: (linc) RNAs make proteins! [J].
Cohen, Stephen M. .
EMBO JOURNAL, 2014, 33 (09) :937-938
[12]   Comprehensive Annotation of Physcomitrella patens Small RNA Loci Reveals That the Heterochromatic Short Interfering RNA Pathway Is Largely Conserved in Land Plants [J].
Coruh, Ceyda ;
Cho, Sung Hyun ;
Shahid, Saima ;
Liu, Qikun ;
Wierzbicki, Andrzej ;
Axtella, Michael J. .
PLANT CELL, 2015, 27 (08) :2148-2162
[13]   Seeing the forest for the trees: annotating small RNA producing genes in plants [J].
Coruh, Ceyda ;
Shahid, Saima ;
Axtell, Michael J. .
CURRENT OPINION IN PLANT BIOLOGY, 2014, 18 :87-95
[14]   Antisense COOLAIR mediates the coordinated switching of chromatin states at FLC during vernalization [J].
Csorba, Tibor ;
Questa, Julia I. ;
Sun, Qianwen ;
Dean, Caroline .
PROCEEDINGS OF THE NATIONAL ACADEMY OF SCIENCES OF THE UNITED STATES OF AMERICA, 2014, 111 (45) :16160-16165
[15]   Unique functionality of 22-nt miRNAs in triggering RDR6-dependent siRNA biogenesis from target transcripts in Arabidopsis [J].
Cuperus, Josh T. ;
Carbonell, Alberto ;
Fahlgren, Noah ;
Garcia-Ruiz, Hernan ;
Burke, Russell T. ;
Takeda, Atsushi ;
Sullivan, Christopher M. ;
Gilbert, Sunny D. ;
Montgomery, Taiowa A. ;
Carrington, James C. .
NATURE STRUCTURAL & MOLECULAR BIOLOGY, 2010, 17 (08) :997-U111
[16]   Genome-wide identification and testing of superior reference genes for transcript normalization in Arabidopsis [J].
Czechowski, T ;
Stitt, M ;
Altmann, T ;
Udvardi, MK ;
Scheible, WR .
PLANT PHYSIOLOGY, 2005, 139 (01) :5-17
[17]   COP1 - A REGULATORY LOCUS INVOLVED IN LIGHT-CONTROLLED DEVELOPMENT AND GENE-EXPRESSION IN ARABIDOPSIS [J].
DENG, XW ;
CASPAR, T ;
QUAIL, PH .
GENES & DEVELOPMENT, 1991, 5 (07) :1172-1182
[18]   Nonsense-Mediated Decay of Alternative Precursor mRNA Splicing Variants Is a Major Determinant of the Arabidopsis Steady State Transcriptome [J].
Drechsel, Gabriele ;
Kahles, Andre ;
Kesarwani, Anil K. ;
Stauffer, Eva ;
Behr, Jonas ;
Drewe, Philipp ;
Raetsch, Gunnar ;
Wachtera, Andreas .
PLANT CELL, 2013, 25 (10) :3726-3742
[19]   Quantitative measures for the management and comparison of annotated genomes [J].
Eilbeck, Karen ;
Moore, Barry ;
Holt, Carson ;
Yandell, Mark .
BMC BIOINFORMATICS, 2009, 10
[20]   Human housekeeping genes, revisited [J].
Eisenberg, Eli ;
Levanon, Erez Y. .
TRENDS IN GENETICS, 2013, 29 (10) :569-574