PubMed Health⌕ Search

PubMed · 14764576

Ontologizing gene-expression microarray data: characterizing clusters with Gene Ontology.

Abstract

An XML-based Java application is described that provides a function-oriented overview of the results of cluster analysis of gene-expression microarray data based on Gene Ontology terms and associations. The application generates one HTML page with listings of the frequencies of explicit and implicit Gene Ontology annotations for each cluster, and separate, linked pages with listings of explicit annotations for each gene in a cluster.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Peter N Robinson, Andreas Wollstein, Ulrike Böhme, Brad Beattie. 2004-02-05. Ontologizing gene-expression microarray data: characterizing clusters with Gene Ontology.. https://doi.org/10.1093/bioinformatics%2Fbth040

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Clustering individuals using INMTD: a novel versatile multi-view embedding framework integrating omics and imaging data.

MOTIVATION: Combining omics and images can lead to a more comprehensive clustering of individuals than classic single-view approaches. Among the various approaches for multi-view clustering, nonnegative matrix tri-factorization (NMTF) and nonnegative Tucker decomposition (NTD) are advantageous in learning low-rank embeddings with promising interpretability. Besides, there is a need to handle unwanted drivers of clusterings (i.e. confounders). RESULTS: In this work, we introduce a novel multi-view clustering method based on NMTF and NTD, named INMTD, which integrates omics and 3D imaging data to derive unconfounded subgroups of individuals. According to the adjusted Rand index, INMTD outperformed other clustering methods on a synthetic dataset with known clusters. In the application to real-life facial-genomic data, INMTD generated biologically relevant embeddings for individuals, genetics, and facial morphology. By removing confounded embedding vectors, we derived an unconfounded clustering with better internal and external quality; the genetic and facial annotations of each derived subgroup highlighted distinctive characteristics. In conclusion, INMTD can effectively integrate omics data and 3D images for unconfounded clustering with biologically meaningful interpretation. AVAILABILITY AND IMPLEMENTATION: INMTD is freely available at https://github.com/ZuqiLi/INMTD.

Cluster Analysis↗

Functional grouping of yeast genes via biclustering microarray data.

Biclustering algorithm on Gibbs sampling strategy is a recruit in the field of the analysis of gene expression data of microarray experiments. Its feasibility and validity still need to be researched not only for synthetic datasets but also for real datasets. Here we investigated a biclustering algorithm on a microarray dataset of Yeast genome through building a database for storing microarray datasets and MIPS data, and running the scripts on Matlab platform to discover gene patterns. In contrast with standard clusterings that reveal genes behaving similarly over all the conditions, biclustering groups genes over only a subset of conditions for which those genes have a sharp probability distribution. It has the key advantage of providing a transparent probabilistic interpretation of the biclusters. Its basic strategy of Gibbs sampling does not suffer from the problem of local minima that often characterizes expectation maximization, so that the patterns should be more global and accurate. Also we tested it with the known explanation of genes in MIPS, objectively to demonstrate the effectiveness and deficiencies of biclustering approach, and the functions of a few unknown ORFs in some bicluster can be deduced in the present research. In addition, the result of similarity searching in Blast-Search can be an assistant evidence for its effectivity.

Cluster Analysis↗