SAliBASE: A Database of Simulated Protein Alignments

被引:2
|
作者
Pervez, Muhammad Tariq [1 ]
Shah, Hayat Ali [2 ]
Babar, Masroor Ellahi [3 ]
Naveed, Nasir [2 ]
Shoaib, Muhammad [4 ]
机构
[1] Virtual Univ Pakistan, Dept Bioinformat & Computat Biol, 54 Lawrence Rd, Lahore 54000, Pakistan
[2] Virtual Univ Pakistan, Dept Comp Sci, Lahore, Pakistan
[3] Virtual Univ Pakistan, Dept Biotechnol, Lahore, Pakistan
[4] UET, Dept Comp Sci & Engn, Lahore, Pakistan
来源
关键词
SAliBASE; simulated alignment; true alignment; MULTIPLE SEQUENCE ALIGNMENT;
D O I
10.1177/1176934318821080
中图分类号
Q [生物科学];
学科分类号
07 ; 0710 ; 09 ;
摘要
Simulated alignments are alternatives to manually constructed multiple sequence alignments for evaluating performance of multiple sequence alignment tools. The importance of simulated sequences is recognized because their true evolutionary history is known, which is very helpful for reconstructing accurate phylogenetic trees and alignments. However, generating simulated alignments require expertise to use bioinformatics tools and consume several hours for reconstructing even a few hundreds of simulated sequences. It becomes a tedious job for an end user who needs a few datasets of variety of simulated sequences. Currently, there is no databank available which may help researchers to download simulated sequences/alignments for their study. Major focus of our study was to develop a database of simulated protein sequences (SAliBASE) based on different varying parameters such as insertion rate, deletion rate, sequence length, number of sequences, and indel size. Each dataset has corresponding alignment as well. This repository is very useful for evaluating multiple alignment methods.
引用
收藏
页数:4
相关论文
共 50 条
  • [1] PROTEIN DATABASE SEARCHES FOR MULTIPLE ALIGNMENTS
    ALTSCHUL, SF
    LIPMAN, DJ
    PROCEEDINGS OF THE NATIONAL ACADEMY OF SCIENCES OF THE UNITED STATES OF AMERICA, 1990, 87 (14) : 5509 - 5513
  • [2] DBAli:: a database of protein structure alignments
    Martí-Renom, MA
    Ilyin, VA
    Sali, A
    BIOINFORMATICS, 2001, 17 (08) : 746 - 747
  • [3] DATABASE OF PROTEIN-SEQUENCE ALIGNMENTS
    BARKER, WC
    GEORGE, DG
    SRINIVASARAO, GY
    YEH, LS
    FASEB JOURNAL, 1992, 6 (01): : A348 - A348
  • [4] DMAPS: a database of multiple alignments for protein structures
    Guda, Chittibabu
    Pal, Lipika R.
    Shindyalov, Ilya N.
    NUCLEIC ACIDS RESEARCH, 2006, 34 : D273 - D276
  • [5] The HSSP database of protein structure sequence alignments
    Schneider, R
    Sander, C
    NUCLEIC ACIDS RESEARCH, 1996, 24 (01) : 201 - 205
  • [6] PALI: a database of alignments and phylogeny of homologous protein structures
    Sujatha, S
    Balaji, S
    Srinivasan, N
    BIOINFORMATICS, 2001, 17 (04) : 375 - 376
  • [7] PIR-ALN: a database of protein sequence alignments
    Srinivasarao, GY
    Yeh, LSL
    Marzec, CR
    Orcutt, BC
    Barker, WC
    BIOINFORMATICS, 1999, 15 (05) : 382 - 390
  • [8] The HSSP database of protein structure-sequence alignments
    Schneider, R
    deDaruvar, A
    Sander, C
    NUCLEIC ACIDS RESEARCH, 1997, 25 (01) : 226 - 230
  • [9] Database of protein sequence alignments: PIR-ALN
    Srinivasarao, GY
    Yeh, LSL
    Marzec, CR
    Orcutt, BC
    Barker, WC
    Pfeiffer, F
    NUCLEIC ACIDS RESEARCH, 1999, 27 (01) : 284 - 285
  • [10] THE HSSP DATABASE OF PROTEIN-STRUCTURE SEQUENCE ALIGNMENTS
    SANDER, C
    SCHNEIDER, R
    NUCLEIC ACIDS RESEARCH, 1994, 22 (17) : 3597 - 3599