Publication:

A Bayesian Framework for the Classification of Microbial Gene Activity States

Loading...
Thumbnail Image

Open/View Files

Date

2016

Published Version

Journal Title

Journal ISSN

Volume Title

Publisher

Frontiers Media S.A.
The Harvard community has made this article openly available. Please share how this access benefits you.

Research Projects

Organizational Units

Journal Issue

Citation

Disselkoen, C., B. Greco, K. Cook, K. Koch, R. Lerebours, C. Viss, J. Cape, et al. 2016. “A Bayesian Framework for the Classification of Microbial Gene Activity States.” Frontiers in Microbiology 7 (1): 1191. doi:10.3389/fmicb.2016.01191. http://dx.doi.org/10.3389/fmicb.2016.01191.

Abstract

Numerous methods for classifying gene activity states based on gene expression data have been proposed for use in downstream applications, such as incorporating transcriptomics data into metabolic models in order to improve resulting flux predictions. These methods often attempt to classify gene activity for each gene in each experimental condition as belonging to one of two states: active (the gene product is part of an active cellular mechanism) or inactive (the cellular mechanism is not active). These existing methods of classifying gene activity states suffer from multiple limitations, including enforcing unrealistic constraints on the overall proportions of active and inactive genes, failing to leverage a priori knowledge of gene co-regulation, failing to account for differences between genes, and failing to provide statistically meaningful confidence estimates. We propose a flexible Bayesian approach to classifying gene activity states based on a Gaussian mixture model. The model integrates genome-wide transcriptomics data from multiple conditions and information about gene co-regulation to provide activity state confidence estimates for each gene in each condition. We compare the performance of our novel method to existing methods on both simulated data and real data from 907 E. coli gene expression arrays, as well as a comparison with experimentally measured flux values in 29 conditions, demonstrating that our method provides more consistent and accurate results than existing methods across a variety of metrics.

Description

Research Data

Keywords

Methods, metabolic modeling, gene expression, bacteria, gene activity, Bayesian model

Terms of Use

This article is made available under the terms and conditions applicable to Other Posted Material (LAA), as set forth at Terms of Service

Endorsement

Review

Supplemented By

Related Stories