A Bioinformatics Pipeline for Whole Exome Sequencing: Overview of the Processing and Steps from Raw Data to Downstream Analysis | Zendy

Narendra Kumar Meena | Zendy; Praveen Mathur | Zendy; Krishna Mohan Medicherla | Zendy; Prashanth Suravajhala | Zendy

Open Access

A Bioinformatics Pipeline for Whole Exome Sequencing: Overview of the Processing and Steps from Raw Data to Downstream Analysis

Author(s) -

Narendra Kumar Meena,

Praveen Mathur,

Krishna Mohan Medicherla,

Prashanth Suravajhala

Publication year - 2018

Publication title -

bio-protocol

Language(s) - English

Resource type - Journals

ISSN - 2331-8325

DOI - 10.21769/bioprotoc.2805

Subject(s) - pipeline (software) , exome sequencing , exome , benchmarking , computer science , computational biology , 1000 genomes project , dna sequencing , data mining , data science , bioinformatics , biology , single nucleotide polymorphism , genetics , mutation , gene , genotype , programming language , marketing , business

[Abstract] Recent advances in Next Generation Sequencing (NGS) technologies have given an impetus to find causality for rare genetic disorders. Since 2005 and aftermath of the human genome project, efforts have been made to understand the rare variants of genetic disorders. Benchmarking the bioinformatics pipeline for whole exome sequencing (WES) has always been a challenge. In this protocol, we discuss detailed steps from quality check to analysis of the variants using a WES pipeline comparing them with reposited public NGS data and survey different techniques, algorithms and software tools used during each step. We observed that variant calling performed on exome and whole genome datasets have different metrics generated when compared to variant callers, GATK and VarScan with different parameters. Furthermore, we found that VarScan with strict parameters could recover 80-85% of high quality GATK SNPs with decreased sensitivity from NGS data. We believe our protocol in the form of pipeline can be used by researchers interested in performing WES analysis for genetic diseases and any clinical phenotypes.

The content you want is available to Zendy users.

Already have an account? Click here to sign in.

Having issues? You can contact us here

Accelerating Research