A Bayesian Framework for the Classification of Microbial Gene Activity States.

Journal Article (Journal Article)

Numerous methods for classifying gene activity states based on gene expression data have been proposed for use in downstream applications, such as incorporating transcriptomics data into metabolic models in order to improve resulting flux predictions. These methods often attempt to classify gene activity for each gene in each experimental condition as belonging to one of two states: active (the gene product is part of an active cellular mechanism) or inactive (the cellular mechanism is not active). These existing methods of classifying gene activity states suffer from multiple limitations, including enforcing unrealistic constraints on the overall proportions of active and inactive genes, failing to leverage a priori knowledge of gene co-regulation, failing to account for differences between genes, and failing to provide statistically meaningful confidence estimates. We propose a flexible Bayesian approach to classifying gene activity states based on a Gaussian mixture model. The model integrates genome-wide transcriptomics data from multiple conditions and information about gene co-regulation to provide activity state confidence estimates for each gene in each condition. We compare the performance of our novel method to existing methods on both simulated data and real data from 907 E. coli gene expression arrays, as well as a comparison with experimentally measured flux values in 29 conditions, demonstrating that our method provides more consistent and accurate results than existing methods across a variety of metrics.

Full Text

Duke Authors

Cited Authors

  • Disselkoen, C; Greco, B; Cook, K; Koch, K; Lerebours, R; Viss, C; Cape, J; Held, E; Ashenafi, Y; Fischer, K; Acosta, A; Cunningham, M; Best, AA; DeJongh, M; Tintle, N

Published Date

  • January 2016

Published In

Volume / Issue

  • 7 /

Start / End Page

  • 1191 -

PubMed ID

  • 27555837

Pubmed Central ID

  • PMC4977825

Electronic International Standard Serial Number (EISSN)

  • 1664-302X

International Standard Serial Number (ISSN)

  • 1664-302X

Digital Object Identifier (DOI)

  • 10.3389/fmicb.2016.01191


  • eng