Back to Search Start Over

An integrative Bayesian Dirichlet-multinomial regression model for the analysis of taxonomic abundances in microbiome data.

Authors :
Wadsworth WD
Argiento R
Guindani M
Galloway-Pena J
Shelburne SA
Vannucci M
Source :
BMC bioinformatics [BMC Bioinformatics] 2017 Feb 08; Vol. 18 (1), pp. 94. Date of Electronic Publication: 2017 Feb 08.
Publication Year :
2017

Abstract

Background: The Human Microbiome has been variously associated with the immune-regulatory mechanisms involved in the prevention or development of many non-infectious human diseases such as autoimmunity, allergy and cancer. Integrative approaches which aim at associating the composition of the human microbiome with other available information, such as clinical covariates and environmental predictors, are paramount to develop a more complete understanding of the role of microbiome in disease development.<br />Results: In this manuscript, we propose a Bayesian Dirichlet-Multinomial regression model which uses spike-and-slab priors for the selection of significant associations between a set of available covariates and taxa from a microbiome abundance table. The approach allows straightforward incorporation of the covariates through a log-linear regression parametrization of the parameters of the Dirichlet-Multinomial likelihood. Inference is conducted through a Markov Chain Monte Carlo algorithm, and selection of the significant covariates is based upon the assessment of posterior probabilities of inclusions and the thresholding of the Bayesian false discovery rate. We design a simulation study to evaluate the performance of the proposed method, and then apply our model on a publicly available dataset obtained from the Human Microbiome Project which associates taxa abundances with KEGG orthology pathways. The method is implemented in specifically developed R code, which has been made publicly available.<br />Conclusions: Our method compares favorably in simulations to several recently proposed approaches for similarly structured data, in terms of increased accuracy and reduced false positive as well as false negative rates. In the application to the data from the Human Microbiome Project, a close evaluation of the biological significance of our findings confirms existing associations in the literature.

Details

Language :
English
ISSN :
1471-2105
Volume :
18
Issue :
1
Database :
MEDLINE
Journal :
BMC bioinformatics
Publication Type :
Academic Journal
Accession number :
28178947
Full Text :
https://doi.org/10.1186/s12859-017-1516-0