Clustering 16S Rrna for OTU Prediction: a Method of Unsupervised Bayesian Clustering.

Xiaolin Hao,Rui Jiang,Ting Chen
DOI: https://doi.org/10.1093/bioinformatics/btq725
IF: 5.8
2011-01-01
Bioinformatics
Abstract:Motivation: With the advancements of next-generation sequencing technology, it is now possible to study samples directly obtained from the environment. Particularly, 16S rRNA gene sequences have been frequently used to profile the diversity of organisms in a sample. However, such studies are still taxed to determine both the number of operational taxonomic units (OTUs) and their relative abundance in a sample.Results: To address these challenges, we propose an unsupervised Bayesian clustering method termed Clustering 16S rRNA for OTU Prediction (CROP). CROP can find clusters based on the natural organization of data without setting a hard cut-off threshold (3%/5%) as required by hierarchical clustering methods. By applying our method to several datasets, we demonstrate that CROP is robust against sequencing errors and that it produces more accurate results than conventional hierarchical clustering methods.
What problem does this paper attempt to address?