The influences of bioinformatics tools and reference databases in analyzing the human oral microbial community

Maria A. Sierra, Qianhao Li, Smruti Pushalkar, Bidisha Paul, Tito A. Sandoval, Angela R. Kamer, Patricia Corby, Yuqi Guo, Ryan Richard Ruff, Alexander V. Alekseyenko, Xin Li, Deepak Saxena

Research output: Contribution to journalArticle

Abstract

There is currently no criterion to select appropriate bioinformatics tools and reference databases for analysis of 16S rRNA amplicon data in the human oral microbiome. Our study aims to determine the influence of multiple tools and reference databases on α-diversity measurements and β-diversity comparisons analyzing the human oral microbiome. We compared the results of taxonomical classification by Greengenes, the Human Oral Microbiome Database (HOMD), National Center for Biotechnology Information (NCBI) 16S, SILVA, and the Ribosomal Database Project (RDP) using Quantitative Insights Into Microbial Ecology (QIIME) and the Divisive Amplicon Denoising Algorithm (DADA2). There were 15 phyla present in all of the analyses, four phyla exclusive to certain databases, and different numbers of genera were identified in each database. Common genera found in the oral microbiome, such as Veillonella, Rothia, and Prevotella, are annotated by all databases; however, less common genera, such as Bulleidia and Paludibacter, are only annotated by large databases, such as Greengenes. Our results indicate that using different reference databases in 16S rRNA amplicon data analysis could lead to different taxonomic compositions, especially at genus level. There are a variety of databases available, but there are no defined criteria for data curation and validation of annotations, which can affect the accuracy and reproducibility of results, making it difficult to compare data across studies.

Original languageEnglish (US)
Article number878
Pages (from-to)1-12
Number of pages12
JournalGenes
Volume11
Issue number8
DOIs
StatePublished - Aug 2020

Keywords

  • 16S rRNA
  • DADA2
  • Databases
  • Greengenes
  • HOMD
  • NCBI
  • QIIME
  • RDP
  • SILVA

ASJC Scopus subject areas

  • Genetics
  • Genetics(clinical)

Fingerprint Dive into the research topics of 'The influences of bioinformatics tools and reference databases in analyzing the human oral microbial community'. Together they form a unique fingerprint.

  • Cite this

    Sierra, M. A., Li, Q., Pushalkar, S., Paul, B., Sandoval, T. A., Kamer, A. R., Corby, P., Guo, Y., Ruff, R. R., Alekseyenko, A. V., Li, X., & Saxena, D. (2020). The influences of bioinformatics tools and reference databases in analyzing the human oral microbial community. Genes, 11(8), 1-12. [878]. https://doi.org/10.3390/genes11080878