Unknown

Dataset Information

0

VCPA: genomic variant calling pipeline and data management tool for Alzheimer's Disease Sequencing Project.


ABSTRACT: SUMMARY:We report VCPA, our SNP/Indel Variant Calling Pipeline and data management tool used for the analysis of whole genome and exome sequencing (WGS/WES) for the Alzheimer's Disease Sequencing Project. VCPA consists of two independent but linkable components: pipeline and tracking database. The pipeline, implemented using the Workflow Description Language and fully optimized for the Amazon elastic compute cloud environment, includes steps from aligning raw sequence reads to variant calling using GATK. The tracking database allows users to view job running status in real time and visualize >100 quality metrics per genome. VCPA is functionally equivalent to the CCDG/TOPMed pipeline. Users can use the pipeline and the dockerized database to process large WGS/WES datasets on Amazon cloud with minimal configuration. AVAILABILITY AND IMPLEMENTATION:VCPA is released under the MIT license and is available for academic and nonprofit use for free. The pipeline source code and step-by-step instructions are available from the National Institute on Aging Genetics of Alzheimer's Disease Data Storage Site (http://www.niagads.org/VCPA). SUPPLEMENTARY INFORMATION:Supplementary data are available at Bioinformatics online.

SUBMITTER: Leung YY 

PROVIDER: S-EPMC6513159 | biostudies-literature | 2019 May

REPOSITORIES: biostudies-literature

altmetric image

Publications

VCPA: genomic variant calling pipeline and data management tool for Alzheimer's Disease Sequencing Project.

Leung Yuk Yee YY   Valladares Otto O   Chou Yi-Fan YF   Lin Han-Jen HJ   Kuzma Amanda B AB   Cantwell Laura L   Qu Liming L   Gangadharan Prabhakaran P   Salerno William J WJ   Schellenberg Gerard D GD   Wang Li-San LS  

Bioinformatics (Oxford, England) 20190501 10


<h4>Summary</h4>We report VCPA, our SNP/Indel Variant Calling Pipeline and data management tool used for the analysis of whole genome and exome sequencing (WGS/WES) for the Alzheimer's Disease Sequencing Project. VCPA consists of two independent but linkable components: pipeline and tracking database. The pipeline, implemented using the Workflow Description Language and fully optimized for the Amazon elastic compute cloud environment, includes steps from aligning raw sequence reads to variant ca  ...[more]

Similar Datasets

| S-EPMC6020218 | biostudies-literature
| S-EPMC5427492 | biostudies-literature
| S-EPMC6756534 | biostudies-literature
| S-EPMC3201884 | biostudies-literature
| S-EPMC7178392 | biostudies-literature
| S-EPMC6722845 | biostudies-literature
| S-EPMC8637281 | biostudies-literature
| S-EPMC5324109 | biostudies-literature
| S-EPMC10794290 | biostudies-literature
| S-EPMC6397097 | biostudies-literature