PubMed Health⌕ Search

PubMed · 10329133

The relationship between protein structure and function: a comprehensive survey with application to the yeast genome.

Abstract

For most proteins in the genome databases, function is predicted via sequence comparison. In spite of the popularity of this approach, the extent to which it can be reliably applied is unknown. We address this issue by systematically investigating the relationship between protein function and structure. We focus initially on enzymes functionally classified by the Enzyme Commission (EC) and relate these to by structurally classified domains the SCOP database. We find that the major SCOP fold classes have different propensities to carry out certain broad categories of functions. For instance, alpha/beta folds are disproportionately associated with enzymes, especially transferases and hydrolases, and all-alpha and small folds with non-enzymes, while alpha+beta folds have an equal tendency either way. These observations for the database overall are largely true for specific genomes. We focus, in particular, on yeast, analyzing it with many classifications in addition to SCOP and EC (i.e. COGs, CATH, MIPS), and find clear tendencies for fold-function association, across a broad spectrum of functions. Analysis with the COGs scheme also suggests that the functions of the most ancient proteins are more evenly distributed among different structural classes than those of more modern ones. For the database overall, we identify the most versatile functions, i.e. those that are associated with the most folds, and the most versatile folds, associated with the most functions. The two most versatile enzymatic functions (hydro-lyases and O-glycosyl glucosidases) are associated with seven folds each. The five most versatile folds (TIM-barrel, Rossmann, ferredoxin, alpha-beta hydrolase, and P-loop NTP hydrolase) are all mixed alpha-beta structures. They stand out as generic scaffolds, accommodating from six to as many as 16 functions (for the exceptional TIM-barrel). At the conclusion of our analysis we are able to construct a graph giving the chance that a functional annotation can be reliably transferred at different degrees of sequence and structural similarity. Supplemental information is available from http://bioinfo.mbb.yale.edu/genome/foldfunc++ +.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

H Hegyi, M Gerstein. 1999-04-23. The relationship between protein structure and function: a comprehensive survey with application to the yeast genome.. https://doi.org/10.1006/jmbi.1999.2661

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

How not to be seen: predicting unseen enzyme functions using contrastive learning.

MOTIVATION: Predicting enzyme function from its sequence is still an unsolved problem in the life sciences. Moreover, with the explosion of annotated genome data, we are inundated with potential enzymatic sequences that have not yet been biochemically characterized. While it is not possible to assign a not-yet-existing label to such a sequence, there is high value in placing the sequence as accurately as possible in known function space. Doing so can help provide more accurate falsifiable hypotheses for experimentalists wishing to characterize enzymes from specific functional families. RESULTS: Here we present a contrastive learning algorithm for predicting enzyme function from sequence. Our method, EnzPlacer, predicts the third, second, and first EC numbers for a protein whose fourth EC number is not in the training corpus. This novel prediction mechanism accurately places a protein sequence within a narrowed-down functional context, even if the precise function remains unknown. AVAILABILITY AND IMPLEMENTATION: EnzPlacer and data is available at https://github.com/drxiangma/EnzPlacer under a GPL3 license.

Enzymes↗

Searchlight on domains.

In this issue of Structure, examine in detail the functions of selected domains within proteins both when they are alone and when in combination with others. Domain function is relevant to molecular evolution and to annotation of proteins known only by sequence.

Enzymes↗

Alpha-secondary isotope effects as probes of "tunneling-ready" configurations in enzymatic H-tunneling: insight from environmentally coupled tunneling models.

Using alpha-secondary kinetic isotope effects (2 degrees KIEs) in conjunction with primary (1 degrees ) KIEs, we have investigated the mechanism of environmentally coupled hydrogen tunneling in the reductive half-reactions of two homologous flavoenzymes, morphinone reductase (MR) and pentaerythritol tetranitrate reductase (PETNR). We find exalted 2 degrees KIEs (1.17-1.18) for both enzymes, consistent with hydrogen tunneling. These 2 degrees KIEs, unlike 1 degrees KIEs, are independent of promoting motions-a nonequilibrium pre-organization of cofactor and active site residues that is required to bring the reactants into a "tunneling-ready" configuration. That these 2 degrees KIEs are identical suggests the geometries of the "tunneling-ready" configurations in both enzymes are indistinguishable, despite the fact that MR, but not PETNR, has a clearly temperature-dependent 1 degrees KIE. The work emphasizes the benefit of combining studies of 1 degrees and 2 degrees KIEs to report on pre-organization and local geometries within the context of contemporary environmentally coupled frameworks for H-tunneling.

Enzymes↗