ggmsa: a visual exploration tool for multiple sequence alignment and associated data

被引:98
|
作者
Zhou, Lang
Feng, Tingze
Xu, Shuangbin
Gao, Fangluan [1 ]
Lam, Tommy T. [2 ]
Wang, Qianwen [3 ]
Wu, Tianzhi [4 ]
Huang, Huina
Zhan, Li
Li, Lin
Guan, Yi [5 ]
Dai, Zehan [6 ]
Yu, Guangchuang [6 ]
机构
[1] Fujian Agr & Forestry Univ, Inst Plant Virol, Fuzhou, Peoples R China
[2] Univ Hong Kong, Sch Publ Hlth, Hong Kong, Peoples R China
[3] Southern Med Univ, Sch Basic Med Sci, Dept Bioinformat, Guangzhou, Peoples R China
[4] Southern Med Univ, Dept Bioinformat, Guangzhou, Peoples R China
[5] Univ Hong Kong, State Key Lab Emerging Infect Dis, Hong Kong, Peoples R China
[6] Southern Med Univ, Guangzhou, Peoples R China
关键词
multiple sequence alignment; sequence bundle; sequence recombination; phylogeny; VISUALIZATION; RNA; RESIDUES;
D O I
10.1093/bib/bbac222
中图分类号
Q5 [生物化学];
学科分类号
071010 ; 081704 ;
摘要
The identification of the conserved and variable regions in the multiple sequence alignment (MSA) is critical to accelerating the process of understanding the function of genes. MSA visualizations allow us to transform sequence features into understandable visual representations. As the sequence-structure-function relationship gains increasing attention in molecular biology studies, the simple display of nucleotide or protein sequence alignment is not satisfied. A more scalable visualization is required to broaden the scope of sequence investigation. Here we present ggmsa, an R package for mining comprehensive sequence features and integrating the associated data of MSA by a variety of display methods. To uncover sequence conservation patterns, variations and recombination at the site level, sequence bundles, sequence logos, stacked sequence alignment and comparative plots are implemented. ggmsa supports integrating the correlation of MSA sequences and their phenotypes, as well as other traits such as ancestral sequences, molecular structures, molecular functions and expression levels. We also design a new visualization method for genome alignments in multiple alignment format to explore the pattern of within and between species variation. Combining these visual representations with prime knowledge, ggmsa assists researchers in discovering MSA and making decisions. The ggmsa package is open-source software released under the Artistic-2.0 license, and it is freely available on Bioconductor (https://bioconductor.org/packages/ggmsa) and Github (https://github.com/YuLab-SMU/ggmsa).
引用
收藏
页数:12
相关论文
共 50 条
  • [21] Profile HMM based Multiple Sequence Alignment for DNA Sequences
    Mulia, Sudipta
    Mishra, Debahuti
    Jena, Tanushree
    INTERNATIONAL CONFERENCE ON MODELLING OPTIMIZATION AND COMPUTING, 2012, 38 : 1783 - 1787
  • [22] A multiple objective evolutionary algorithm for multiple sequence alignment
    Seeluangsawat, Pasut
    Chongstitvatana, Prabhas
    GECCO 2005: Genetic and Evolutionary Computation Conference, Vols 1 and 2, 2005, : 477 - 478
  • [23] Small design from big alignment: engineering proteins with multiple sequence alignment as the starting point
    Wang, Tianwen
    Liang, Chen
    Hou, Yajing
    Zheng, Mengyuan
    Xu, Hongju
    An, Yafei
    Xiao, Sa
    Liu, Lu
    Lian, Shuaibin
    BIOTECHNOLOGY LETTERS, 2020, 42 (08) : 1305 - 1315
  • [24] Visual exploration of microbiome data
    Kuntal, Bhusan K.
    Mande, Sharmila S.
    JOURNAL OF BIOSCIENCES, 2019, 44 (05)
  • [25] Multiple Sequence Alignment of the M Protein in SARS-Associated and Other Known Coronaviruses
    史定华
    周晖杰
    王斌宾
    顾燕红
    王翼飞
    Advances in Manufacturing, 2003, (02) : 118 - 123
  • [26] Optimization of consistency-based multiple sequence alignment using Big Data technologies
    Jordi Lladós
    Fernando Cores
    Fernando Guirado
    The Journal of Supercomputing, 2019, 75 : 1310 - 1322
  • [27] Visual exploration of microbiome data
    Bhusan K. Kuntal
    Sharmila S. Mande
    Journal of Biosciences, 2019, 44
  • [28] Optimization of consistency-based multiple sequence alignment using Big Data technologies
    Llados, Jordi
    Cores, Fernando
    Guirado, Fernando
    JOURNAL OF SUPERCOMPUTING, 2019, 75 (03) : 1310 - 1322
  • [29] An Algorithm of Multiple Sequence Alignment Based on Consensus Sequence Searched by Simulated Annealing and Star Alignment
    Yao, Dengfeng
    Jiang, Minghu
    You, Xu
    Abulizi, Abudoukelimu
    Hou, Renkui
    2015 INTERNATIONAL SYMPOSIUM ON BIOELECTRONICS AND BIOINFORMATICS (ISBB), 2015, : 3 - 6
  • [30] A NEW GENETIC ALGORITHM FOR MULTIPLE SEQUENCE ALIGNMENT
    Narimani, Zahra
    Beigy, Hamid
    Abolhassani, Hassan
    INTERNATIONAL JOURNAL OF COMPUTATIONAL INTELLIGENCE AND APPLICATIONS, 2012, 11 (04)