PubMed Health⌕ Search

PubMed · 11248037

Reverse engineering the (beta/alpha )8 barrel fold.

Abstract

The (beta/alpha)(8) barrel is the most commonly occurring fold among protein catalysts. To lay a groundwork for engineering novel barrel proteins, we investigated the amino acid sequence restrictions at 182 structural positions of the prototypical (beta/alpha)(8) barrel enzyme triosephosphate isomerase. Using combinatorial mutagenesis and functional selection, we find that turn sequences, alpha-helix capping and stop motifs, and residues that pack the interface between beta-strands and alpha-helices are highly mutable. Conversely, any mutation of residues in the central core of the beta-barrel, beta-strand stop motifs, and a single buried salt bridge between amino acids R189 and D227 substantially reduces catalytic activity. Four positions are effectively immutable: conservative single substitutions at these four positions prevent the mutant protein from complementing a triosephosphate isomerase knockout in Escherichia coli. At 142 of the 182 positions, mutation to at least one amino acid of a seven-letter amino acid alphabet produces a triosephosphate isomerase with wild-type activity. Consequently, it seems likely that (beta/alpha)(8) barrel structures can be encoded with a subset of the 20 amino acids. Such simplification would greatly decrease the computational burden of (beta/alpha)(8) barrel design.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

J A Silverman, R Balakrishnan, P B Harbury. 2001-03-13. Reverse engineering the (beta/alpha )8 barrel fold.. https://doi.org/10.1073/pnas.041613598

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Base-pair resolution conservation data improves cell type specific sequence-to-expression prediction.

MOTIVATION: Genomic sequence-to-activity models can decipher gene regulatory mechanisms and predict the functional impact of regulatory variants. However, current models struggle to integrate information from sequences outside promoters, especially information from cell type specific regulatory elements. RESULTS: Here, we propose incorporating base-pair resolution evolutionary conservation data into genomic sequence-to-expression predictors. We explore two training strategies-training from scratch or fine-tuning an existing sequence-only model with additional conservation input. We find that in both cases, base-pair resolution conservation data improves cell type specific sequence-to-expression prediction, with training from scratch yielding the greatest benefit. The improvement in cell type specific expression prediction can be attributed in part to the fact that models trained on sequence and conservation data learn to better recognize cell type specific regulatory elements than models trained on sequence alone. AVAILABILITY: Code is available at https://github.com/ni-lab/basenji-phyloP.

Conserved Sequence↗

Lift&Add-rapid and robust addition of new species to alignments of conserved non-coding sequences.

MOTIVATION: Identifying sequence constraint across long evolutionary distances is a powerful method for the discovery of functional genomic sequences, especially putative non-coding elements. Conserved elements have been a mainstay of comparative genomic research, and can be further investigated for species-specific sequence acceleration to dissect the genetic basis of trait evolution. The conclusions of these comparative genomic studies are contingent on the number and range of species included in this phylogenetic analysis. However, while the number of metazoan genomes sequences is increasing rapidly, adding new genomes to existing whole-genome alignments remains computationally expensive. RESULTS: Here, we present a bioinformatic workflow, Lift&Add, that enables conserved elements, coding or non-coding, to be rapidly mapped to new genomes ("Lift") and subsequently be added to pre-existing multiple species alignments ("Add"), thus providing an avenue for easy exploration of these putative functional elements. Focusing here on a group of species that has been largely under-represented in genomic comparisons, the marsupials, we demonstrate the intuition behind this workflow and provide an example comparative genomic analysis that can be performed. IMPLEMENTATION AND AVAILABILITY: Lift&Add is implemented as a series of scripts in Snakemake and bash, which can be downloaded from https://github.com/navyashukladr/Lift_and_Add.

Conserved Sequence↗

Analysis of the structure of human telomerase RNA in vivo.

Telomerase is a ribonucleoprotein reverse transcriptase that synthesises telomeric DNA. The RNA component of telomerase acts as a template for telomere synthesis and binds the reverse transcriptase. In this study, we have performed in vivo and in vitro structural analyses of human telomerase RNA (hTR). In vivo mapping experiments showed that the 5'-terminal template domain of hTR folds into a long hairpin structure, in which the template sequence occupies a readily accessible position. Intriguingly, neither in vivo nor in vitro mapping of hTR confirmed formation of a stable 'pseudoknot' helix, suggesting that this functionally essential long range interaction is formed only temporarily. In vitro control mappings demonstrated that the 5'-terminal template domain of hTR cannot fold correctly in the absence of cellular protein factors. The 3'-terminal domain of hTR, both in vivo and in vitro, folds into the previously predicted box H/ACA snoRNA-like 'hairpin-hinge-hairpin-tail' structure. Finally, comparison of the in vivo and in vitro modification patterns of hTR revealed several regions that might be directly involved in binding of telomerase reverse transcriptase or other telomerase proteins.

Conserved Sequence↗