Rich annotation of DNA sequencing variants by leveraging the Ensembl Variant Effect Predictor with plugins.
Where this comes from
- Record sourced from PubMed, PMID 24626529.
- Also identified by DOI 10.1093/bib/bbu008 and PMC identifier 6283364.
- No licence information is recorded for this record.
- Because redistribution is not established, this page shows the abstract only. Follow the links below for the full text.
Abstract
High-throughput DNA sequencing has become a mainstay for the discovery of genomic variants that may cause disease or affect phenotype. A next-generation sequencing pipeline typically identifies thousands of variants in each sample. A particular challenge is the annotation of each variant in a way that is useful to downstream consumers of the data, such as clinical sequencing centers or researchers. These users may require that all data storage and analysis remain on secure local servers to protect patient confidentiality or intellectual property, may have unique and changing needs to draw on a variety of annotation data sets and may prefer not to rely on closed-source applications beyond their control. Here we describe scalable methods for using the plugin capability of the Ensembl Variant Effect Predictor to enrich its basic set of variant annotations with additional data on genes, function, conservation, expression, diseases, pathways and protein structure, and describe an extensible framework for easily adding additional custom data sets.
Medical subject headings
- High-Throughput Nucleotide Sequencing
- Molecular Sequence Annotation
- Sequence Analysis, DNA