Accuracy of microbial community diversity estimated by closed- and open-reference OTUs
- PMID: 29018622
- PMCID: PMC5631090
- DOI: 10.7717/peerj.3889
Accuracy of microbial community diversity estimated by closed- and open-reference OTUs
Abstract
Next-generation sequencing of 16S ribosomal RNA is widely used to survey microbial communities. Sequences are typically assigned to Operational Taxonomic Units (OTUs). Closed- and open-reference OTU assignment matches reads to a reference database at 97% identity (closed), then clusters unmatched reads using a de novo method (open). Implementations of these methods in the QIIME package were tested on several mock community datasets with 20 strains using different sequencing technologies and primers. Richness (number of reported OTUs) was often greatly exaggerated, with hundreds or thousands of OTUs generated on Illumina datasets. Between-sample diversity was also found to be highly exaggerated in many cases, with weighted Jaccard distances between identical mock samples often close to one, indicating very low similarity. Non-overlapping hyper-variable regions in 70% of species were assigned to different OTUs. On mock communities with Illumina V4 reads, 56% to 88% of predicted genus names were false positives. Biological inferences obtained using these methods are therefore not reliable.
Keywords: Alpha diversity; Beta diversity; Closed-reference; OTU; Open-reference; QIIME.
Conflict of interest statement
The author is the author of software tools that provide alternatives to the methods evaluated in this manuscript.
Figures
References
-
- Bergey DH. Bergey’s manual of systematic bacteriology. Springer; London: 2001.
-
- Caporaso JG, Kuczynski J, Stombaugh J, Bittinger K, Bushman FD, Costello EK, Fierer N, Peña AG, Goodrich JK, Gordon JI, Huttley GA, Kelley ST, Knights D, Koenig JE, Ley RE, Lozupone CA, McDonald D, Muegge BD, Pirrung M, Reeder J, Sevinsky JR, Turnbaugh PJ, Walters WA, Widmann J, Yatsunenko T, Zaneveld J, Knight R. QIIME allows analysis of high-throughput community sequencing data. Nature Methods. 2010;7:335–336. doi: 10.1038/nmeth.f.303. - DOI - PMC - PubMed
LinkOut - more resources
Full Text Sources
Other Literature Sources
