PubMed Health⌕ Search

PubMed · 15561146

A domain interaction map based on phylogenetic profiling.

Abstract

Phylogenetic profiling is a well established method for predicting functional relations and physical interactions between proteins. We present a new method for finding such relations based on phylogenetic profiling of conserved domains rather than proteins, avoiding computationally expensive all versus all sequence comparisons among genomes. The resulting domain interaction map (DIMA) can be explored directly or mapped to a genome of interest. We demonstrate that the performance of DIMA is comparable to that of classical phylogenetic profiling and its predictions often yield information that cannot be detected by profiling of entire protein chains. We provide a list of novel domain associations predicted by our method.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Philipp Pagel, Philip Wong, Dmitrij Frishman. 2004-12-10. A domain interaction map based on phylogenetic profiling.. https://doi.org/10.1016/j.jmb.2004.10.019

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Clustering individuals using INMTD: a novel versatile multi-view embedding framework integrating omics and imaging data.

MOTIVATION: Combining omics and images can lead to a more comprehensive clustering of individuals than classic single-view approaches. Among the various approaches for multi-view clustering, nonnegative matrix tri-factorization (NMTF) and nonnegative Tucker decomposition (NTD) are advantageous in learning low-rank embeddings with promising interpretability. Besides, there is a need to handle unwanted drivers of clusterings (i.e. confounders). RESULTS: In this work, we introduce a novel multi-view clustering method based on NMTF and NTD, named INMTD, which integrates omics and 3D imaging data to derive unconfounded subgroups of individuals. According to the adjusted Rand index, INMTD outperformed other clustering methods on a synthetic dataset with known clusters. In the application to real-life facial-genomic data, INMTD generated biologically relevant embeddings for individuals, genetics, and facial morphology. By removing confounded embedding vectors, we derived an unconfounded clustering with better internal and external quality; the genetic and facial annotations of each derived subgroup highlighted distinctive characteristics. In conclusion, INMTD can effectively integrate omics data and 3D images for unconfounded clustering with biologically meaningful interpretation. AVAILABILITY AND IMPLEMENTATION: INMTD is freely available at https://github.com/ZuqiLi/INMTD.

Cluster Analysis↗

Multiscale detection of localized anomalous structure in aggregate disease incidence data.

We present a modelling framework for detection of potentially anomalous structure in aggregate spatial disease incidence data in a manner sensitive to localization at multiple scales and/or positions. The key technical contribution is the re-casting of the components of a multiscale disease mapping methodology, recently introduced by the authors in an earlier paper, into a form appropriate for hypothesis testing. In particular, we describe how hypotheses of spatially clustered variations in disease incidence may be linked in one-to-one correspondence with collections of hypotheses on the values of certain multiscale parameters associated with a user-defined hierarchy of nested partitions of an overall spatial region. A Bayesian hypothesis testing methodology is developed in the context of a standard Poisson measurement model, over the collection of possible multiscale hypotheses. We discuss the specification of hyper parameters and prior distributions on the space of models. The methodology is illustrated on both simulated and real data.

Cluster Analysis↗