Negative binomial mixed models for analyzing microbiome count data
Open Access
- 3 January 2017
- journal article
- research article
- Published by Springer Science and Business Media LLC in BMC Bioinformatics
- Vol. 18 (1), 1-10
- https://doi.org/10.1186/s12859-016-1441-7
Abstract
Recent advances in next-generation sequencing (NGS) technology enable researchers to collect a large volume of metagenomic sequencing data. These data provide valuable resources for investigating interactions between the microbiome and host environmental/clinical factors. In addition to the well-known properties of microbiome count measurements, for example, varied total sequence reads across samples, over-dispersion and zero-inflation, microbiome studies usually collect samples with hierarchical structures, which introduce correlation among the samples and thus further complicate the analysis and interpretation of microbiome count data. In this article, we propose negative binomial mixed models (NBMMs) for detecting the association between the microbiome and host environmental/clinical factors for correlated microbiome count data. Although having not dealt with zero-inflation, the proposed mixed-effects models account for correlation among the samples by incorporating random effects into the commonly used fixed-effects negative binomial model, and can efficiently handle over-dispersion and varying total reads. We have developed a flexible and efficient IWLS (Iterative Weighted Least Squares) algorithm to fit the proposed NBMMs by taking advantage of the standard procedure for fitting the linear mixed models. We evaluate and demonstrate the proposed method via extensive simulation studies and the application to mouse gut microbiome data. The results show that the proposed method has desirable properties and outperform the previously used methods in terms of both empirical power and Type I error. The method has been incorporated into the freely available R package BhGLM ( http://www.ssg.uab.edu/bhglm/ and http://github.com/abbyyan3/BhGLM ), providing a useful tool for analyzing microbiome data.Keywords
Funding Information
- National Institutes of Health (R01GM069430)
- National Institutes of Health (R03DE024198)
- National Institutes of Health (DK087346)
- National Natural Science Foundation of China (81573253, 31571291)
This publication has 65 references indexed in Scilit:
- Dietary Fat Content and Fiber Type Modulate Hind Gut Microbial Community and Metabolic Markers in the PigPLOS ONE, 2013
- Correlation between body mass index and gut concentrations of Lactobacillus reuteri, Bifidobacterium animalis, Methanobrevibacter smithii and Escherichia coliInternational Journal of Obesity, 2013
- Human gut microbiome viewed across age and geographyNature, 2012
- The human microbiome: at the interface of health and diseaseNature Reviews Genetics, 2012
- Human-Associated Microbial Signatures: Examining Their Predictive ValueCell Host & Microbe, 2011
- The Future of microbial metagenomics (or is ignorance bliss?)The ISME Journal, 2011
- Metagenomics: Facts and Artifacts, and Computational ChallengesJournal of Computer Science and Technology, 2010
- A core gut microbiome in obese and lean twinsNature, 2008
- Worlds within worlds: evolution of the vertebrate gut microbiotaNature Reviews Microbiology, 2008
- An obesity-associated gut microbiome with increased capacity for energy harvestNature, 2006