High quality SNP calling using Illumina data at shallow coverage
High quality SNP calling using Illumina data at shallow coverage
复制标题
DOI:
10.1093/bioinformatics/btq092
复制
发表时间:
2010-04-15
期刊:
影响因子:
5.8
通讯作者:
Jones, Steven J. M.
中科院分区:
文献类型:
--
作者:
Malhis, Nawar;Jones, Steven J. M.
Motivation: Detection of single nucleotide polymorphisms (SNPs) has been a major application in processing second generation sequencing (SGS) data. In principle, SNPs are called on single base differences between a reference genome and a sequence generated from SGS short reads of a sample genome. However, this exercise is far from trivial; several parameters related to sequencing quality, and/or reference genome properties, play essential effect on the accuracy of called SNPs especially at shallow coverage data. In this work, we present Slider II, an alignment and SNP calling approach that demonstrates improved algorithmic approaches enabling larger number of called SNPs with lower false positive rate. In addition to the regular alignment and SNP calling, as an optional feature, Slider II is capable of utilizing information about known SNPs of a target genome, as priors, in the alignment and SNPs calling to enhance it's capability of detecting these known SNPs and novel SNPs and mutations in their vicinity.