Lineage-Based Identification of Cellular States and Expression Programs

DSpace/Manakin Repository

Lineage-Based Identification of Cellular States and Expression Programs

Citable link to this page


Title: Lineage-Based Identification of Cellular States and Expression Programs
Author: Jaakkola, Tommi; Mazzoni, Esteban O.; Wichterle, Hynek; Hashimoto, Tatsunori Benjamin; Sherwood, Richard Irving; Gifford, David Kenneth

Note: Order does not necessarily reflect citation order of authors.

Citation: Hashimoto, Tatsunori, Tommi Jaakkola, Richard Sherwood, Esteban O. Mazzoni, Hynek Wichterle, and David Gifford. 2012. Lineage-based identification of cellular states and expression programs. Bioinformatics 28(12): i250-i257.
Full Text & Related Files:
Abstract: Summary: We present a method, LineageProgram, that uses the developmental lineage relationship of observed gene expression measurements to improve the learning of developmentally relevant cellular states and expression programs. We find that incorporating lineage information allows us to significantly improve both the predictive power and interpretability of expression programs that are derived from expression measurements from in vitro differentiation experiments. The lineage tree of a differentiation experiment is a tree graph whose nodes describe all of the unique expression states in the input expression measurements, and edges describe the experimental perturbations applied to cells. Our method, LineageProgram, is based on a log-linear model with parameters that reflect changes along the lineage tree. Regularization with L1 that based methods controls the parameters in three distinct ways: the number of genes change between two cellular states, the number of unique cellular states, and the number of underlying factors responsible for changes in cell state. The model is estimated with proximal operators to quickly discover a small number of key cell states and gene sets. Comparisons with existing factorization, techniques, such as singular value decomposition and non-negative matrix factorization show that our method provides higher predictive power in held, out tests while inducing sparse and biologically relevant gene sets.
Published Version: doi:10.1093/bioinformatics/bts204
Other Sources:
Terms of Use: This article is made available under the terms and conditions applicable to Other Posted Material, as set forth at
Citable link to this page:
Downloads of this work:

Show full Dublin Core record

This item appears in the following Collection(s)


Search DASH

Advanced Search