A practical method to detect SNVs and indels from whole genome and exome sequencing data
- PMID: 23831772
- PMCID: PMC3703611
- DOI: 10.1038/srep02161
A practical method to detect SNVs and indels from whole genome and exome sequencing data
Abstract
The recent development of massively parallel sequencing technology has allowed the creation of comprehensive catalogs of genetic variation. However, due to the relatively high sequencing error rate for short read sequence data, sophisticated analysis methods are required to obtain high-quality variant calls. Here, we developed a probabilistic multinomial method for the detection of single nucleotide variants (SNVs) as well as short insertions and deletions (indels) in whole genome sequencing (WGS) and whole exome sequencing (WES) data for single sample calling. Evaluation with DNA genotyping arrays revealed a concordance rate of 99.98% for WGS calls and 99.99% for WES calls. Sanger sequencing of the discordant calls determined the false positive and false negative rates for the WGS (0.0068% and 0.17%) and WES (0.0036% and 0.0084%) datasets. Furthermore, short indels were identified with high accuracy (WGS: 94.7%, WES: 97.3%). We believe our method can contribute to the greater understanding of human diseases.
Figures
References
-
- Rusk N. & Kiermer V. Primer: Sequencing--the next generation. Nat Methods 5, 15 (2008). - PubMed
-
- Metzker M. L. Sequencing technologies - the next generation. Nat Rev Genet 11, 31–46 (2010). - PubMed
-
- Mardis E. R. Next-generation DNA sequencing methods. Annu Rev Genomics Hum Genet 9, 387–402 (2008). - PubMed
Publication types
MeSH terms
LinkOut - more resources
Full Text Sources
Other Literature Sources
