PubMed Health⌕ Search

PubMed · 15033362

Quantifying structure-function uncertainty: a graph theoretical exploration into the origins and limitations of protein annotation.

Abstract

Since the advent of investigations into structural genomics, research has focused on correctly identifying domain boundaries, as well as domain similarities and differences in the context of their evolutionary relationships. As the science of structural genomics ramps up adding more and more information into the databanks, questions about the accuracy and completeness of our classification and annotation systems appear on the forefront of this research. A central question of paramount importance is how structural similarity relates to functional similarity. Here, we begin to rigorously and quantitatively answer these questions by first exploring the consensus between the most common protein domain structure annotation databases CATH, SCOP and FSSP. Each of these databases explores the evolutionary relationships between protein domains using a combination of automatic and manual, structural and functional, continuous and discrete similarity measures. In order to examine the issue of consensus thoroughly, we build a generalized graph out of each of these databases and hierarchically cluster these graphs at interval thresholds. We then employ a distance measure to find regions of greatest overlap. Using this procedure we were able not only to enumerate the level of consensus between the different annotation systems, but also to define the graph-theoretical origins behind the annotation schema of class, family and superfamily by observing that the same thresholds that define the best consensus regions between FSSP, SCOP and CATH correspond to distinct, non-random phase-transitions in the structure comparison graph itself. To investigate the correspondence in divergence between structure and function further, we introduce a measure of functional entropy that calculates divergence in function space. First, we use this measure to calculate the general correlation between structural homology and functional proximity. We extend this analysis further by quantitatively calculating the average amount of functional information gained from our understanding of structural distance and the corollary inherent uncertainty that represents the theoretical limit of our ability to infer function from structural similarity. Finally we show how our measure of functional "entropy" translates into a more intuitive concept of functional annotation into similarity EC classes.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Boris E Shakhnovich, J Max Harvey. 2004-04-02. Quantifying structure-function uncertainty: a graph theoretical exploration into the origins and limitations of protein annotation.. https://doi.org/10.1016/j.jmb.2004.02.009

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

"Pocket dendrimers" as nanoscale receptors for bimolecular guest accommodation.

A new series of dendrimer receptors was prepared by combining a (tetraphenylporphinato)zinc(II) core and benzyl ether type dendritic substituents. Since one direction of the (tetraphenylporphinato)zinc(II) was not substituted by a dendritic residue, the resulting unsymmetrical dendrimers have "pockets" available for access of external substrates. Molecular modeling, NMR measurements, and zinc-coordination experiments revealed that the third-generation dendrimer of this type exhibited characteristic inclusion of coordinative pyridine guests. When diamidopyridine moiety was introduced into the dendrimer pocket, a thymine derivative was bound through complementary hydrogen bonding. Two different kinds of substrates, pyridine and thymine derivatives, were simultaneously accommodated in the nanoscale pocket and bimolecular guest accommodation was realized with the designed dendrimer receptor.

Biochemical Phenomena↗

Distribution of selenium in different biochemical fractions and raw darkening degree of potato (Solanum tuberosum L.) tubers supplemented with selenate.

Effects of Se fertilization on potato processing quality, possible changes in Se concentration and form in tubers during storage, and retransfer of Se from seed tubers were examined. Potato plants were grown at five selenate (SeO4(2-)) concentrations. Tubers were harvested 16 weeks after planting and were stored at 3-4 degrees C prior to analysis. The results showed that the Se concentration did not decrease during storage for 1-12 months. In tubers, 49-65% of total Se was allocated in protein fraction, which is less than found in plant leaves in a previous study. The next-generation tubers produced by the Se-enriched seed tubers had increased Se concentrations, which evidenced the relocation of Se from the seed tubers. At low levels, Se improved the processing quality of potato tubers by diminishing and retarding their raw darkening. The value of Se-enriched potato tubers as a Se source in the human diet was discussed.

Biochemical Phenomena↗

Human development V: biochemistry unable to explain the emergence of biological form (morphogenesis) and therefore a new principle as source of biological information is needed.

Today's biomedicine builds on the conviction that biochemistry can explain the creation of the body, its anatomy and physiology. Unfortunately there are still deep mysteries strangely "fighting back" when we try to define and understand the organism and its creation in the ontogenesis as emerging from biochemistry. In analysing this from a theoretical perspective using a mathematical model focusing on the noise in complex chemical systems we argue that evolving biological structure cannot in principle be a product of chemistry. In this paper we go through the chemical gradient model and argue that this is not able to explain the ontogenesis. We discuss the used gradients as information carriers in chemical self-organizing systems and argue that by use of the "Turing structures" we are only able to modelling the mostly simple biological systems. The bio-chemical model is only able to model simple organization but not to explain the complexity of biological phenomena. We conclude that we seemingly have presented a formal proof (a NO-GO theorem) that the self-organizing chemical systems that are using chemical gradients are not able to explain complex biological matters as the ontogenesis. We need a fundamentally new, information-carrying principle to understand biological information and biological order.

Biochemical Phenomena↗