RCSB Protein Data Bank (RCSB.org): delivery of experimentally-determined PDB structures alongside one million computed structure models of proteins from artificial intelligence/machine learning

被引:424
作者
Burley, Stephen K. [1 ,2 ,3 ,4 ,5 ]
Bhikadiya, Charmi [5 ]
Bi, Chunxiao [5 ]
Bittrich, Sebastian [5 ]
Chao, Henry [1 ,2 ]
Chen, Li [1 ,2 ]
Craig, Paul A. [6 ]
Crichlow, Gregg, V [1 ,2 ]
Dalenberg, Kenneth [1 ,2 ]
Duarte, Jose M. [5 ]
Dutta, Shuchismita [1 ,2 ,3 ]
Fayazi, Maryam [1 ,2 ]
Feng, Zukang [1 ,2 ]
Flatt, Justin W. [1 ,2 ]
Ganesan, Sai [7 ]
Ghosh, Sutapa [1 ,2 ]
Goodsell, David S. [1 ,2 ,3 ,8 ]
Green, Rachel Kramer [1 ,2 ]
Guranovic, Vladimir [1 ,2 ]
Henry, Jeremy [5 ]
Hudson, Brian P. [1 ,2 ]
Khokhriakov, Igor [5 ]
Lawson, Catherine L. [1 ,2 ]
Liang, Yuhe [1 ,2 ]
Lowe, Robert [1 ,2 ]
Peisach, Ezra [1 ,2 ]
Persikova, Irina [1 ,2 ]
Piehl, Dennis W. [1 ,2 ]
Rose, Yana [5 ]
Sali, Andrej
Segura, Joan [5 ]
Sekharan, Monica [1 ,2 ]
Shao, Chenghua [1 ,2 ]
Vallat, Brinda [1 ,2 ]
Voigt, Maria [1 ,2 ]
Webb, Ben
Westbrook, John D. [1 ,2 ,3 ]
Whetstone, Shamara [1 ,2 ]
Young, Jasmine Y. [1 ,2 ]
Zalevsky, Arthur
Zardecki, Christine [1 ,2 ]
机构
[1] Rutgers State Univ, Res Collaboratory Struct Bioinformat Prot Data Ban, Piscataway, NJ 08854 USA
[2] Rutgers State Univ, Inst Quantitat Biomed, Piscataway, NJ 08854 USA
[3] Rutgers Canc Inst New Jersey, New Brunswick, NJ 08901 USA
[4] Rutgers State Univ, Dept Chem & Chem Biol, Piscataway, NJ 08854 USA
[5] Univ Calif San Diego, San Diego Supercomp Ctr, Res Collaboratory Struct Bioinformat Prot Data Ban, La Jolla, CA 92093 USA
[6] Rochester Inst Technol, Sch Chem & Mat Sci, Rochester, NY 14623 USA
[7] Univ Calif San Francisco, Quantitat Biosci Inst, Dept Bioengn & Therapeut Sci, Dept Pharmaceut Chem,Res Collaboratory Struct Bioi, San Francisco, CA 94158 USA
[8] Scripps Res Inst, Dept Integrat Struct & Computat Biol, La Jolla, CA 92037 USA
基金
美国国家卫生研究院; 美国国家科学基金会;
关键词
STRUCTURE PREDICTION; BIOLOGICAL MACROMOLECULES; RESOURCE; INFORMATION; VALIDATION; BIOTECHNOLOGY; VISUALIZATION; ANNOTATION; DEPOSITION; FAMILIES;
D O I
10.1093/nar/gkac1077
中图分类号
Q5 [生物化学]; Q7 [分子生物学];
学科分类号
071010 ; 081704 ;
摘要
The Research Collaboratory for Structural Bioinformatics Protein Data Bank (RCSB PDB), founding member of the Worldwide Protein Data Bank (wwPDB), is the US data center for the open-access PDB archive. As wwPDB-designated Archive Keeper, RCSB PDB is also responsible for PDB data security. Annually, RCSB PDB serves >10 000 depositors of three-dimensional (3D) biostructures working on all permanently inhabited continents. RCSB PDB delivers data from its research-focused RCSB.org web portal to many millions of PDB data consumers based in virtually every United Nations-recognized country, territory, etc. This Database Issue contribution describes upgrades to the research-focused RCSB.org web portal that created a one-stop-shop for open access to similar to 200 000 experimentally-determined PDB structures of biological macromolecules alongside >1 000 000 incorporated Computed Structure Models (CSMs) predicted using artificial intelligence/machine learning methods. RCSB.org is a 'living data resource.' Every PDB structure and CSM is integrated weekly with related functional annotations from external biodata resources, providing up-to-date information for the entire corpus of 3D biostructure data freely available from RCSB.org with no usage limitations. Within RCSB.org, PDB structures and the CSMs are clearly identified as to their provenance and reliability. Both are fully searchable, and can be analyzed and visualized using the full complement of RCSB.org web portal capabilities.
引用
收藏
页码:D488 / D508
页数:21
相关论文
共 128 条
[41]   The Pfam protein families database: towards a more sustainable future [J].
Finn, Robert D. ;
Coggill, Penelope ;
Eberhardt, Ruth Y. ;
Eddy, Sean R. ;
Mistry, Jaina ;
Mitchell, Alex L. ;
Potter, Simon C. ;
Punta, Marco ;
Qureshi, Matloob ;
Sangrador-Vegas, Amaia ;
Salazar, Gustavo A. ;
Tate, John ;
Bateman, Alex .
NUCLEIC ACIDS RESEARCH, 2016, 44 (D1) :D279-D285
[42]  
Fitzgerald P.M.D., 2005, International tables for crystallography G definition and exchange of crystallographic data, P295
[43]   The RESID database of protein modifications as a resource and annotation tool [J].
Garavelli, JS .
PROTEOMICS, 2004, 4 (06) :1527-1533
[44]   The ChEMBL database in 2017 [J].
Gaulton, Anna ;
Hersey, Anne ;
Nowotka, Michal ;
Bento, A. Patricia ;
Chambers, Jon ;
Mendez, David ;
Mutowo, Prudence ;
Atkinson, Francis ;
Bellis, Louisa J. ;
Cibrian-Uhalte, Elena ;
Davies, Mark ;
Dedman, Nathan ;
Karlsson, Anneli ;
Magarinos, Maria Paula ;
Overington, John P. ;
Papadatos, George ;
Smit, Ines ;
Leach, Andrew R. .
NUCLEIC ACIDS RESEARCH, 2017, 45 (D1) :D945-D954
[45]   BindingDB in 2015: A public database for medicinal chemistry, computational chemistry and systems pharmacology [J].
Gilson, Michael K. ;
Liu, Tiqing ;
Baitaluk, Michael ;
Nicola, George ;
Hwang, Linda ;
Chong, Jenny .
NUCLEIC ACIDS RESEARCH, 2016, 44 (D1) :D1045-D1053
[46]   RCSB Protein Data Bank resources for structure-facilitated design of mRNA vaccines for existing and emerging viral pathogens [J].
Goodsell, David S. ;
Burley, Stephen K. .
STRUCTURE, 2022, 30 (01) :55-+
[47]   RCSB Protein Data Bank: Enabling biomedical research and drug discovery [J].
Goodsell, David S. ;
Zardecki, Christine ;
Di Costanzo, Luigi ;
Duarte, Jose M. ;
Hudson, Brian P. ;
Persikova, Irina ;
Segura, Joan ;
Shao, Chenghua ;
Voigt, Maria ;
Westbrook, John D. ;
Young, Jasmine Y. ;
Burley, Stephen K. .
PROTEIN SCIENCE, 2020, 29 (01) :52-65
[48]   Validation of Structures in the Protein Data Bank [J].
Gore, Swanand ;
Garcia, Eduardo Sanz ;
Hendrickx, Pieter M. S. ;
Gutmanas, Aleksandras ;
Westbrook, John D. ;
Yang, Huanwang ;
Feng, Zukang ;
Baskaran, Kumaran ;
Berrisford, John M. ;
Hudson, Brian P. ;
Ikegawa, Yasuyo ;
Kobayashi, Naohiro ;
Lawson, Catherine L. ;
Mading, Steve ;
Mak, Lora ;
Mukhopadhyay, Abhik ;
Oldfield, Thomas J. ;
Patwardhan, Ardan ;
Peisach, Ezra ;
Sahni, Gaurav ;
Sekharan, Monica R. ;
Sen, Sanchayita ;
Shao, Chenghua ;
Smart, Oliver S. ;
Ulrich, Eldon L. ;
Yamashita, Reiko ;
Quesada, Martha ;
Young, Jasmine Y. ;
Nakamura, Haruki ;
Markley, John L. ;
Berman, Helen M. ;
Burley, Stephen K. ;
Velankar, Sameer ;
Kleywegt, Gerard J. .
STRUCTURE, 2017, 25 (12) :1916-1927
[49]   The Cambridge Structural Database [J].
Groom, Colin R. ;
Bruno, Ian J. ;
Lightfoot, Matthew P. ;
Ward, Suzanna C. .
ACTA CRYSTALLOGRAPHICA SECTION B-STRUCTURAL SCIENCE CRYSTAL ENGINEERING AND MATERIALS, 2016, 72 :171-179
[50]   Real time structural search of the Protein Data Bank [J].
Guzenko, Dmytro ;
Burley, Stephen K. ;
Duarte, Jose M. .
PLOS COMPUTATIONAL BIOLOGY, 2020, 16 (07)