Comparison of sequencing based CNV discovery methods using monozygotic twin quartets.
Comparison of sequencing based CNV discovery methods using monozygotic twin quartets.
复制标题
DOI:
10.1371/journal.pone.0122287
复制
发表时间:
2015
期刊:
影响因子:
3.7
通讯作者:
Dubé MP
中科院分区:
文献类型:
--
作者:
Legault MA;Girard S;Lemieux Perreault LP;Rouleau GA;Dubé MP
The advent of high throughput sequencing methods breeds an important amount of technical challenges. Among those is the one raised by the discovery of copy-number variations (CNVs) using whole-genome sequencing data. CNVs are genomic structural variations defined as a variation in the number of copies of a large genomic fragment, usually more than one kilobase. Here, we aim to compare different CNV calling methods in order to assess their ability to consistently identify CNVs by comparison of the calls in 9 quartets of identical twin pairs. The use of monozygotic twins provides a means of estimating the error rate of each algorithm by observing CNVs that are inconsistently called when considering the rules of Mendelian inheritance and the assumption of an identical genome between twins. The similarity between the calls from the different tools and the advantage of combining call sets were also considered. ERDS and CNVnator obtained the best performance when considering the inherited CNV rate with a mean of 0.74 and 0.70, respectively. Venn diagrams were generated to show the agreement between the different algorithms, before and after filtering out familial inconsistencies. This filtering revealed a high number of false positives for CNVer and Breakdancer. A low overall agreement between the methods suggested a high complementarity of the different tools when calling CNVs. The breakpoint sensitivity analysis indicated that CNVnator and ERDS achieved better resolution of CNV borders than the other tools. The highest inherited CNV rate was achieved through the intersection of these two tools (81%). This study showed that ERDS and CNVnator provide good performance on whole genome sequencing data with respect to CNV consistency across families, CNV breakpoint resolution and CNV call specificity. The intersection of the calls from the two tools would be valuable for CNV genotyping pipelines.
登录
查看更多内容
影响因子:
3
作者:
Chu JH;Rogers A;Ionita-Laza I;Darvishi K;Mills RE;Lee C;Raby BA
通讯作者:
Raby BA
影响因子:
9.5
作者:
van de Wiel, Mark A.;Picard, Franck;Ylstra, Bauke
通讯作者:
Ylstra, Bauke
影响因子:
30.8
作者:
Iafrate, AJ;Feuk, L;Lee, C
通讯作者:
Lee, C
影响因子:
56.9
作者:
Sebat, J;Lakshmi, B;Wigler, M
通讯作者:
Wigler, M
影响因子:
5.2
作者:
Kantaputra, Piranit N.;Klopocki, Eva;Mundlos, Stefan
通讯作者:
Mundlos, Stefan