ContextMap 2: fast and accurate context-based RNA-seq mapping

被引:36
作者
Bonfert, Thomas [1 ]
Kirner, Evelyn [1 ]
Csaba, Gergely [1 ]
Zimmer, Ralf [1 ]
Friedel, Caroline C. [1 ]
机构
[1] Univ Munich, Inst Informat, D-80333 Munich, Germany
来源
BMC BIOINFORMATICS | 2015年 / 16卷
关键词
READ ALIGNMENT;
D O I
10.1186/s12859-015-0557-5
中图分类号
Q5 [生物化学];
学科分类号
071010 ; 081704 ;
摘要
Background: Mapping of short sequencing reads is a crucial step in the analysis of RNA sequencing (RNA-seq) data. ContextMap is an RNA-seq mapping algorithm that uses a context-based approach to identify the best alignment for each read and allows parallel mapping against several reference genomes. Results: In this article, we present ContextMap 2, a new and improved version of ContextMap. Its key novel features are: (i) a plug-in structure that allows easily integrating novel short read alignment programs with improved accuracy and runtime; (ii) context-based identification of insertions and deletions (indels); (iii) mapping of reads spanning an arbitrary number of exons and indels. ContextMap 2 using Bowtie, Bowtie 2 or BWA was evaluated on both simulated and real-life data from the recently published RGASP study. Conclusions: We show that ContextMap 2 generally combines similar or higher recall compared to other state-of-the-art approaches with significantly higher precision in read placement and junction and indel prediction. Furthermore, runtime was significantly lower than for the best competing approaches. ContextMap 2 is freely available at http://www.bio.ifi.lmu.de/ContextMap.
引用
收藏
页数:15
相关论文
共 19 条
  • [1] Alamancos GP, 2014, METHODS MOL BIOL, V1126, P357, DOI 10.1007/978-1-62703-980-2_26
  • [2] Bonfert T, 2012, BMC BIOINF S6, V13, P9
  • [3] Mining RNA-Seq Data for Infections and Contaminations
    Bonfert, Thomas
    Csaba, Gergely
    Zimmer, Ralf
    Friedel, Caroline C.
    [J]. PLOS ONE, 2013, 8 (09):
  • [4] Landscape of transcription in human cells
    Djebali, Sarah
    Davis, Carrie A.
    Merkel, Angelika
    Dobin, Alex
    Lassmann, Timo
    Mortazavi, Ali
    Tanzer, Andrea
    Lagarde, Julien
    Lin, Wei
    Schlesinger, Felix
    Xue, Chenghai
    Marinov, Georgi K.
    Khatun, Jainab
    Williams, Brian A.
    Zaleski, Chris
    Rozowsky, Joel
    Roeder, Maik
    Kokocinski, Felix
    Abdelhamid, Rehab F.
    Alioto, Tyler
    Antoshechkin, Igor
    Baer, Michael T.
    Bar, Nadav S.
    Batut, Philippe
    Bell, Kimberly
    Bell, Ian
    Chakrabortty, Sudipto
    Chen, Xian
    Chrast, Jacqueline
    Curado, Joao
    Derrien, Thomas
    Drenkow, Jorg
    Dumais, Erica
    Dumais, Jacqueline
    Duttagupta, Radha
    Falconnet, Emilie
    Fastuca, Meagan
    Fejes-Toth, Kata
    Ferreira, Pedro
    Foissac, Sylvain
    Fullwood, Melissa J.
    Gao, Hui
    Gonzalez, David
    Gordon, Assaf
    Gunawardena, Harsha
    Howald, Cedric
    Jha, Sonali
    Johnson, Rory
    Kapranov, Philipp
    King, Brandon
    [J]. NATURE, 2012, 489 (7414) : 101 - 108
  • [5] STAR: ultrafast universal RNA-seq aligner
    Dobin, Alexander
    Davis, Carrie A.
    Schlesinger, Felix
    Drenkow, Jorg
    Zaleski, Chris
    Jha, Sonali
    Batut, Philippe
    Chaisson, Mark
    Gingeras, Thomas R.
    [J]. BIOINFORMATICS, 2013, 29 (01) : 15 - 21
  • [6] An integrated encyclopedia of DNA elements in the human genome
    Dunham, Ian
    Kundaje, Anshul
    Aldred, Shelley F.
    Collins, Patrick J.
    Davis, CarrieA.
    Doyle, Francis
    Epstein, Charles B.
    Frietze, Seth
    Harrow, Jennifer
    Kaul, Rajinder
    Khatun, Jainab
    Lajoie, Bryan R.
    Landt, Stephen G.
    Lee, Bum-Kyu
    Pauli, Florencia
    Rosenbloom, Kate R.
    Sabo, Peter
    Safi, Alexias
    Sanyal, Amartya
    Shoresh, Noam
    Simon, Jeremy M.
    Song, Lingyun
    Trinklein, Nathan D.
    Altshuler, Robert C.
    Birney, Ewan
    Brown, James B.
    Cheng, Chao
    Djebali, Sarah
    Dong, Xianjun
    Dunham, Ian
    Ernst, Jason
    Furey, Terrence S.
    Gerstein, Mark
    Giardine, Belinda
    Greven, Melissa
    Hardison, Ross C.
    Harris, Robert S.
    Herrero, Javier
    Hoffman, Michael M.
    Iyer, Sowmya
    Kellis, Manolis
    Khatun, Jainab
    Kheradpour, Pouya
    Kundaje, Anshul
    Lassmann, Timo
    Li, Qunhua
    Lin, Xinying
    Marinov, Georgi K.
    Merkel, Angelika
    Mortazavi, Ali
    [J]. NATURE, 2012, 489 (7414) : 57 - 74
  • [7] Engström PG, 2013, NAT METHODS, V10, P1185, DOI [10.1038/nmeth.2722, 10.1038/NMETH.2722]
  • [8] Garber M, 2011, NAT METHODS, V8, P469, DOI [10.1038/NMETH.1613, 10.1038/nmeth.1613]
  • [9] Comparative analysis of RNA-Seq alignment algorithms and the RNA-Seq unified mapper (RUM)
    Grant, Gregory R.
    Farkas, Michael H.
    Pizarro, Angel D.
    Lahens, Nicholas F.
    Schug, Jonathan
    Brunk, Brian P.
    Stoeckert, Christian J.
    Hogenesch, John B.
    Pierce, Eric A.
    [J]. BIOINFORMATICS, 2011, 27 (18) : 2518 - 2528
  • [10] TopHat2: accurate alignment of transcriptomes in the presence of insertions, deletions and gene fusions
    Kim, Daehwan
    Pertea, Geo
    Trapnell, Cole
    Pimentel, Harold
    Kelley, Ryan
    Salzberg, Steven L.
    [J]. GENOME BIOLOGY, 2013, 14 (04):