Assessment of kinship detection using RNA-seq data

被引:9
作者
Blay, Natalia [1 ,2 ,3 ]
Casas, Eduard [1 ,2 ,4 ]
Galvan-Femenia, Ivan [1 ,5 ]
Graffelman, Jan [6 ,7 ]
de Cid, Rafael [1 ,5 ]
Vavouri, Tanya [1 ,2 ]
机构
[1] Germans Trias & Pujol Res Inst PMPPC IGTP, Program Predict & Personalized Med Canc, Badalona 08916, Spain
[2] Univ Autonoma Barcelona, Josep Carreras Leukaemia Res Inst IJC, Campus ICO Germans Trias & Pujol, Badalona 08916, Spain
[3] Univ Oberta Catalunya, Masters Programme Bioinformat & Biostat, Barcelona 08035, Spain
[4] Univ Barcelona, Doctoral Programme Biomed, Barcelona 08007, Spain
[5] Germans Trias & Pujol Res Inst, Genomes Life GCAT Lab Grp, Can Ruti Campus,Cami Escoles S-N, Barcelona 08916, Spain
[6] Univ Politecn Cataluna, Dept Stat & Operat Res, Barcelona 08028, Spain
[7] Univ Washington, Dept Biostat, Seattle, WA 98105 USA
关键词
IDENTIFICATION; RELATEDNESS; INHERITANCE; LIKELIHOODS; EXPRESSION; VARIANTS; FORMAT; TOOL;
D O I
10.1093/nar/gkz776
中图分类号
Q5 [生物化学]; Q7 [分子生物学];
学科分类号
071010 ; 081704 ;
摘要
Analysis of RNA sequencing (RNA-seq) data from related individuals is widely used in clinical and molecular genetics studies. Prediction of kinship from RNA-seq data would be useful for confirming the expected relationships in family based studies and for highlighting samples from related individuals in case-control or population based studies. Currently, reconstruction of pedigrees is largely based on SNPs or microsatellites, obtained from genotyping arrays, whole genome sequencing and whole exome sequencing. Potential problems with using RNA-seq data for kinship detection are the low proportion of the genome that it covers, the highly skewed coverage of exons of different genes depending on expression level and allele-specific expression. In this study we assess the use of RNA-seq data to detect kinship between individuals, through pairwise identity by descent (IBD) estimates. First, we obtained high quality SNPs after successive filters to minimize the effects due to allelic imbalance as well as errors in sequencing, mapping and genotyping. Then, we used these SNPs to calculate pairwise IBD estimates. By analysing both real and simulated RNA-seq data we show that it is possible to identify up to second degree relationships using RNA-seq data of even low to moderate sequencing depth.
引用
收藏
页数:9
相关论文
共 41 条
[21]   The Sequence Alignment/Map format and SAMtools [J].
Li, Heng ;
Handsaker, Bob ;
Wysoker, Alec ;
Fennell, Tim ;
Ruan, Jue ;
Homer, Nils ;
Marth, Gabor ;
Abecasis, Goncalo ;
Durbin, Richard .
BIOINFORMATICS, 2009, 25 (16) :2078-2079
[22]   Transcriptome Sequencing of a Large Human Family Identifies the Impact of Rare Noncoding Variants [J].
Li, Xin ;
Battle, Alexis ;
Karczewski, Konrad J. ;
Zappala, Zach ;
Knowles, David A. ;
Smith, Kevin S. ;
Kukurba, Kim R. ;
Wu, Eric ;
Simon, Noah ;
Montgomery, Stephen B. .
AMERICAN JOURNAL OF HUMAN GENETICS, 2014, 95 (03) :245-256
[23]   The Genome Analysis Toolkit: A MapReduce framework for analyzing next-generation DNA sequencing data [J].
McKenna, Aaron ;
Hanna, Matthew ;
Banks, Eric ;
Sivachenko, Andrey ;
Cibulskis, Kristian ;
Kernytsky, Andrew ;
Garimella, Kiran ;
Altshuler, David ;
Gabriel, Stacey ;
Daly, Mark ;
DePristo, Mark A. .
GENOME RESEARCH, 2010, 20 (09) :1297-1303
[24]   THE RELATIONSHIP BETWEEN SINGLE PARENT AND PARENT PAIR GENETIC LIKELIHOODS IN GENEALOGY RECONSTRUCTION [J].
MEAGHER, TR ;
THOMPSON, E .
THEORETICAL POPULATION BIOLOGY, 1986, 29 (01) :87-106
[25]   Second Case of HOIP Deficiency Expands Clinical Features and Defines Inflammatory Transcriptome Regulated by LUBAC [J].
Oda, Hirotsugu ;
Beck, David B. ;
Kuehn, Hye Sun ;
Moura, Natalia Sampaio ;
Hoffmann, Patrycja ;
Ibarra, Maria ;
Stoddard, Jennifer ;
Tsai, Wanxia Li ;
Gutierrez-Cruz, Gustavo ;
Gadina, Massimo ;
Rosenzweig, Sergio D. ;
Kastner, Daniel L. ;
Notarangelo, Luigi D. ;
Aksentijevich, Ivona .
FRONTIERS IN IMMUNOLOGY, 2019, 10
[26]   Reliable Identification of Genomic Variants from RNA-Seq Data [J].
Piskol, Robert ;
Ramaswami, Gokul ;
Li, Jin Billy .
AMERICAN JOURNAL OF HUMAN GENETICS, 2013, 93 (04) :641-651
[27]  
Pournelle G. H., 1953, Journal of Mammalogy, V34, P133
[28]   Single-Tissue and Cross-Tissue Heritability of Gene Expression Via Identity-by-Descent in Related or Unrelated Individuals [J].
Price, Alkes L. ;
Helgason, Agnar ;
Thorleifsson, Gudmar ;
McCarroll, Steven A. ;
Kong, Augustine ;
Stefansson, Kari .
PLOS GENETICS, 2011, 7 (02)
[29]   PLINK: A tool set for whole-genome association and population-based linkage analyses [J].
Purcell, Shaun ;
Neale, Benjamin ;
Todd-Brown, Kathe ;
Thomas, Lori ;
Ferreira, Manuel A. R. ;
Bender, David ;
Maller, Julian ;
Sklar, Pamela ;
de Bakker, Paul I. W. ;
Daly, Mark J. ;
Sham, Pak C. .
AMERICAN JOURNAL OF HUMAN GENETICS, 2007, 81 (03) :559-575
[30]   BEDTools: a flexible suite of utilities for comparing genomic features [J].
Quinlan, Aaron R. ;
Hall, Ira M. .
BIOINFORMATICS, 2010, 26 (06) :841-842