A new and updated resource for codon usage tables

被引:181
作者
Athey, John [1 ]
Alexaki, Aikaterini [1 ]
Osipova, Ekaterina [2 ]
Rostovtsev, Alexandre [2 ]
Santana-Quintero, Luis V. [2 ]
Katneni, Upendra [1 ]
Simonyan, Vahan [2 ]
Kimchi-Sarfaty, Chava [1 ]
机构
[1] US FDA, Div Plasma Prot Therapeut, Off Tissue & Adv Therapies, Ctr Biol Evaluat & Res, Silver Spring, MD 20993 USA
[2] US FDA, High Performance Integrated Environm, Ctr Biol Evaluat & Res, Silver Spring, MD USA
关键词
Codon usage bias; Codon optimization; Recombinant protein therapeutics; Translational kinetics; HIGH-LEVEL EXPRESSION; ESCHERICHIA-COLI; GENE-THERAPY; HUMAN INTERLEUKIN-2; TRANSFER-RNAS; BIAS; OPTIMIZATION; EVOLUTION; NUMBER; CONSERVATION;
D O I
10.1186/s12859-017-1793-7
中图分类号
Q5 [生物化学];
学科分类号
071010 ; 081704 ;
摘要
Background: Due to the degeneracy of the genetic code, most amino acids can be encoded by multiple synonymous codons. Synonymous codons naturally occur with different frequencies in different organisms. The choice of codons may affect protein expression, structure, and function. Recombinant gene technologies commonly take advantage of the former effect by implementing a technique termed codon optimization, in which codons are replaced with synonymous ones in order to increase protein expression. This technique relies on the accurate knowledge of codon usage frequencies. Accurately quantifying codon usage bias for different organisms is useful not only for codon optimization, but also for evolutionary and translation studies: phylogenetic relations of organisms, and host-pathogen co-evolution relationships, may be explored through their codon usage similarities. Furthermore, codon usage has been shown to affect protein structure and function through interfering with translation kinetics, and cotranslational protein folding. Results: Despite the obvious need for accurate codon usage tables, currently available resources are either limited in scope, encompassing only organisms from specific domains of life, or greatly outdated. Taking advantage of the exponential growth of GenBank and the creation of NCBI's RefSeq database, we have developed a new database, the High-performance Integrated Virtual Environment-Codon Usage Tables (HIVE-CUTs), to present and analyse codon usage tables for every organism with publicly available sequencing data. Compared to existing databases, this new database is more comprehensive, addresses concerns that limited the accuracy of earlier databases, and provides several new functionalities, such as the ability to view and compare codon usage between individual organisms and across taxonomical clades, through graphical representation or through commonly used indices. In addition, it is being routinely updated to keep up with the continuous flow of new data in GenBank and RefSeq. Conclusion: Given the impact of codon usage bias on recombinant gene technologies, this database will facilitate effective development and review of recombinant drug products and will be instrumental in a wide area of biological research. The database is available at hive. biochemistry.gwu.edu/review/codon.
引用
收藏
页数:10
相关论文
共 65 条
[1]  
AKASHI H, 1994, GENETICS, V136, P927
[2]   PRINCIPLES THAT GOVERN FOLDING OF PROTEIN CHAINS [J].
ANFINSEN, CB .
SCIENCE, 1973, 181 (4096) :223-230
[3]   Codon usage: Nature's roadmap to expression and folding of proteins [J].
Angov, Evelina .
BIOTECHNOLOGY JOURNAL, 2011, 6 (06) :650-659
[4]  
Angov E, 2011, METHODS MOL BIOL, V705, P1, DOI 10.1007/978-1-61737-967-3_1
[5]   Heterologous Protein Expression Is Enhanced by Harmonizing the Codon Usage Frequencies of the Target Gene with those of the Expression Host [J].
Angov, Evelina ;
Hillier, Collette J. ;
Kincaid, Randall L. ;
Lyon, Jeffrey A. .
PLOS ONE, 2008, 3 (05)
[6]   Origin of the 1918 pandemic H1N1 influenza A virus as studied by codon usage patterns and phylogenetic analysis [J].
Anhlan, Darisuren ;
Grundmann, Norbert ;
Makalowski, Wojciech ;
Ludwig, Stephan ;
Scholtissek, Christoph .
RNA, 2011, 17 (01) :64-73
[7]   Overcoming codon bias:: A method for high-level overexpression of Plasmodium and other AT-rich parasite genes in Escherichia coli [J].
Baca, AM ;
Hol, WGJ .
INTERNATIONAL JOURNAL FOR PARASITOLOGY, 2000, 30 (02) :113-118
[8]   Viral adaptation to host: a proteome-based analysis of codon usage and amino acid preferences [J].
Bahir, Iris ;
Fromer, Menachem ;
Prat, Yosef ;
Linial, Michal .
MOLECULAR SYSTEMS BIOLOGY, 2009, 5
[9]   Decoding mechanisms by which silent codon changes influence protein biogenesis and function [J].
Bali, Vedrana ;
Bebok, Zsuzsanna .
INTERNATIONAL JOURNAL OF BIOCHEMISTRY & CELL BIOLOGY, 2015, 64 :58-74
[10]  
Benson DA, 2010, NUCLEIC ACIDS RES, V38, pD46, DOI [10.1093/nar/gkp1024, 10.1093/nar/gkx1094, 10.1093/nar/gkl986, 10.1093/nar/gkw1070, 10.1093/nar/gks1195, 10.1093/nar/gkn723, 10.1093/nar/gkg057, 10.1093/nar/gkr1202, 10.1093/nar/gkq1079]