Home LiteratureArticle Details
PMID: 11473015 Published · ppublish English Journal Article Research Support, Non-U.S. Gov't Research Support, U.S. Gov't, Non-P.H.S. Research Support, U.S. Gov't, P.H.S.

Rich probabilistic models for gene expression.

Bioinformatics (Oxford, England) ·Vol. 17 Suppl 1 ·2001-00-00 ·Pages S243-52

Segal E, Taskar B, Gasch A, Friedman N, Koller D

Abstract

Clustering is commonly used for analyzing gene expression data. Despite their successes, clustering methods suffer from a number of limitations. First, these methods reveal similarities that exist over all of the measurements, while obscuring relationships that exist over only a subset of the data. Second, clustering methods cannot readily incorporate additional types of information, such as clinical data or known attributes of genes. To circumvent these shortcomings, we propose the use of a single coherent probabilistic model, that encompasses much of the rich structure in the genomic expression data, while incorporating additional information such as experiment type, putative binding sites, or functional information. We show how this model can be learned from the data, allowing us to discover patterns in the data and dependencies between the gene expression patterns and additional attributes. The learned model reveals context-specific relationships, that exist only over a subset of the experiments in the dataset. We demonstrate the power of our approach on synthetic data and on two real-world gene expression data sets for yeast. For example, we demonstrate a novel functionality that falls naturally out of our framework: predicting the "cluster" of the array resulting from a gene mutation based only on the gene's expression pattern in the context of other mutations.

MeSH Terms
Algorithms Artificial Intelligence Cluster Analysis Computational Biology Databases, Genetic Gene Expression Profiling/statistics & numerical data Genes, Fungal Genetic Techniques/statistics & numerical data Models, Statistical Saccharomyces cerevisiae/genetics
Authors & Affiliations
5 authors, click to expand affiliations / ORCID
Segal E
Computer Science Department, Stanford University, Stanford 94305, USA. [email protected]
Taskar B
Gasch A
Friedman N
Koller D
Article Info
Journal
Bioinformatics (Oxford, England)
Abbr.
Bioinformatics
ISSN
1367-4803
Published
2001-00-00
Pages
S243-52
Language
English
Region
England
NLM ID
9808944
Subset
IM
Grants
NHLBI NIH HHS · HL07279-23 · United States
Analysis Services
Analysis Services

Contact

No. 2 Wenbo Road, Zhangqiu District, Jinan, Shandong

Qilu Normal University · Genelibs Bioinformatics Lab

750 Shunhua Rd, Jinan

2F, Bldg F, University Science Park

Tel: 0531-88819269

WeChat Official Account

Follow our WeChat subscription account for real-time updates and the latest in medical and biological research.


Business Email

E-mail: [email protected]