Rich annotation of DNA sequencing variants by leveraging the Ensembl Variant Effect Predictor with plugins
Rich annotation of DNA sequencing variants by leveraging the Ensembl Variant Effect Predictor with plugins
复制标题
DOI:
10.1093/bib/bbu008
复制
发表时间:
2015-03-01
影响因子:
9.5
通讯作者:
Nelson, Stanley F.
中科院分区:
文献类型:
--
作者:
Yourshaw, Michael;Taylor, S. Paige;Nelson, Stanley F.
High-throughput DNA sequencing has become a mainstay for the discovery of genomic variants that may cause disease or affect phenotype. A next-generation sequencing pipeline typically identifies thousands of variants in each sample. A particular challenge is the annotation of each variant in a way that is useful to downstream consumers of the data, such as clinical sequencing centers or researchers. These users may require that all data storage and analysis remain on secure local servers to protect patient confidentiality or intellectual property, may have unique and changing needs to draw on a variety of annotation data sets and may prefer not to rely on closed-source applications beyond their control. Here we describe scalable methods for using the plugin capability of the Ensembl Variant Effect Predictor to enrich its basic set of variant annotations with additional data on genes, function, conservation, expression, diseases, pathways and protein structure, and describe an extensible framework for easily adding additional custom data sets.