共 31 条
Prediction of non-coding and antisense RNA genes in Escherichia coli with Gapped Markov Model
被引:29
作者:
Yachie, Nozomu
Numata, Koji
Saito, Rintaro
[1
]
Kanai, Akio
Tomita, Masaru
机构:
[1] Keio Univ, Inst Adv Biosci, Tsuruoka, Yamagata 9970017, Japan
[2] Keio Univ, Inst Adv Biosci, Tsuruoka 9970035, Japan
[3] Keio Univ, Grad Sch Media & Governance, Bioinformat Program, Fujisawa, Kanagawa 2528520, Japan
[4] Keio Univ, Dept Environm Informat, Fujisawa, Kanagawa 2528520, Japan
来源:
关键词:
bioinformatics;
Markov model;
small RNA (sRNA);
sigma70;
promoter;
Rho-independent terminator;
D O I:
10.1016/j.gene.2005.12.034
中图分类号:
Q3 [遗传学];
学科分类号:
071007 ;
090102 ;
摘要:
A new mathematical index was developed to identify and characterize non-coding RNA (ncRNA) genes encoded within the Escherichia coli (E. coli) genome. It was designated the GMMI (Gapped Markov Model Index) and used to evaluate sequence patterns located at the separate positions of consensus sequences, codon biases and/or possible RNA structures on the basis of the Markov model. The GMMI was able to separate a set of known mRNA sequences from a mixture of ncRNAs including tRNAs and rRNAs. Consequently, the GMMI was employed to predict novel ncRNA candidates. At the beginning, possible transcription units were extracted from the E. coli genome using consensus sequences for the sigma70 promoter and the rho-independent terminator. Then, these units were evaluated by using the GMMI. This identified 133 candidate ncRNAs, which contain 29 previously annotated small RNA genes and 46 possible antisense ncRNAs.Furthermore 12 transcripts (including five antisense RNAs) were confirmed according to the expression analysis. These data suggests that the expression of small antisense RNAs might be more common than previously thought in the E. coli genome. (c) 2006 Elsevier B.V. All rights reserved.
引用
收藏
页码:171 / 181
页数:11
相关论文