PubMed Health⌕ Search

PubMed · 15215430

ArrayXPath: mapping and visualizing microarray gene-expression data with integrated biological pathway resources using Scalable Vector Graphics.

Abstract

Biological pathways can provide key information on the organization of biological systems. ArrayXPath (http://www.snubi.org/software/ArrayXPath/) is a web-based service for mapping and visualizing microarray gene-expression data for integrated biological pathway resources using Scalable Vector Graphics (SVG). By integrating major bio-databases and searching pathway resources, ArrayXPath automatically maps different types of identifiers from microarray probes and pathway elements. When one inputs gene-expression clusters, ArrayXPath produces a list of the best matching pathways for each cluster. We applied Fisher's exact test and the false discovery rate (FDR) to evaluate the statistical significance of the association between a cluster and a pathway while correcting the multiple-comparison problem. ArrayXPath produces Javascript-enabled SVGs for web-enabled interactive visualization of pathways integrated with gene-expression profiles.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Hee-Joon Chung, Mingoo Kim, Chan Hee Park, Jihoon Kim, Ju Han Kim. 2004-07-01. ArrayXPath: mapping and visualizing microarray gene-expression data with integrated biological pathway resources using Scalable Vector Graphics.. https://doi.org/10.1093/nar%2Fgkh476

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Clustering individuals using INMTD: a novel versatile multi-view embedding framework integrating omics and imaging data.

MOTIVATION: Combining omics and images can lead to a more comprehensive clustering of individuals than classic single-view approaches. Among the various approaches for multi-view clustering, nonnegative matrix tri-factorization (NMTF) and nonnegative Tucker decomposition (NTD) are advantageous in learning low-rank embeddings with promising interpretability. Besides, there is a need to handle unwanted drivers of clusterings (i.e. confounders). RESULTS: In this work, we introduce a novel multi-view clustering method based on NMTF and NTD, named INMTD, which integrates omics and 3D imaging data to derive unconfounded subgroups of individuals. According to the adjusted Rand index, INMTD outperformed other clustering methods on a synthetic dataset with known clusters. In the application to real-life facial-genomic data, INMTD generated biologically relevant embeddings for individuals, genetics, and facial morphology. By removing confounded embedding vectors, we derived an unconfounded clustering with better internal and external quality; the genetic and facial annotations of each derived subgroup highlighted distinctive characteristics. In conclusion, INMTD can effectively integrate omics data and 3D images for unconfounded clustering with biologically meaningful interpretation. AVAILABILITY AND IMPLEMENTATION: INMTD is freely available at https://github.com/ZuqiLi/INMTD.

Cluster Analysis↗

Rationalising germplasm collections: a case study for wheat.

In total 70 genebank accessions comprising 50 hexaploid, 12 tetraploid and 8 diploid wheats of the Gatersleben collection were selected based on the screening of the passport data for identical cultivar names or accession numbers of the donor genebanks. Twelve potential duplicate groups consisting of three to nine accessions with identical names/numbers were selected and analysed with DNA markers (microsatellites). A bootstrap approach based on re-sampling of both microsatellite markers and alleles within marker loci was used to test for homogeneity. Although several homogeneous groups were identified it became clear that cultivar name identity alone did not allow the determination of duplicates. A combination of SSR-analysis followed by the bootstrap method and database survey considering the botanical classification and other data (origin, growth habit and donor) available is recommended in order to determine duplicates. A procedure for the identification of duplicates and their further handling in ex situ genebanks is discussed.

Cluster Analysis↗