PubMed Health⌕ Search

PubMed · 11287961

Functional proteins from a random-sequence library.

Abstract

Functional primordial proteins presumably originated from random sequences, but it is not known how frequently functional, or even folded, proteins occur in collections of random sequences. Here we have used in vitro selection of messenger RNA displayed proteins, in which each protein is covalently linked through its carboxy terminus to the 3' end of its encoding mRNA, to sample a large number of distinct random sequences. Starting from a library of 6 x 1012 proteins each containing 80 contiguous random amino acids, we selected functional proteins by enriching for those that bind to ATP. This selection yielded four new ATP-binding proteins that appear to be unrelated to each other or to anything found in the current databases of biological proteins. The frequency of occurrence of functional proteins in random-sequence libraries appears to be similar to that observed for equivalent RNA libraries.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

A D Keefe, J W Szostak. 2001-04-05. Functional proteins from a random-sequence library.. https://doi.org/10.1038/35070613

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Common variants at ABCA7, MS4A6A/MS4A4E, EPHA1, CD33 and CD2AP are associated with Alzheimer's disease.

We sought to identify new susceptibility loci for Alzheimer's disease through a staged association study (GERAD+) and by testing suggestive loci reported by the Alzheimer's Disease Genetic Consortium (ADGC) in a companion paper. We undertook a combined analysis of four genome-wide association datasets (stage 1) and identified ten newly associated variants with P ≤ 1 × 10(-5). We tested these variants for association in an independent sample (stage 2). Three SNPs at two loci replicated and showed evidence for association in a further sample (stage 3). Meta-analyses of all data provided compelling evidence that ABCA7 (rs3764650, meta P = 4.5 × 10(-17); including ADGC data, meta P = 5.0 × 10(-21)) and the MS4A gene cluster (rs610932, meta P = 1.8 × 10(-14); including ADGC data, meta P = 1.2 × 10(-16)) are new Alzheimer's disease susceptibility loci. We also found independent evidence for association for three loci reported by the ADGC, which, when combined, showed genome-wide significance: CD2AP (GERAD+, P = 8.0 × 10(-4); including ADGC data, meta P = 8.6 × 10(-9)), CD33 (GERAD+, P = 2.2 × 10(-4); including ADGC data, meta P = 1.6 × 10(-9)) and EPHA1 (GERAD+, P = 3.4 × 10(-4); including ADGC data, meta P = 6.0 × 10(-10)).

ATP-Binding Cassette Transporters↗

BATMAS30: amino acid substitution matrix for alignment of bacterial transporters.

Aligned amino acid sequences of three functionally independent samples of transmembrane (TM) transport proteins have been analyzed. The concept of TM-kernel is proposed as the most probable transmembrane region of a sequence. The average amino acid composition of TM-kernels differs from the published amino acid composition of transmembrane segments. TM-kernels contain more alanines, glycines, and less polar, charged, and aromatic residues in contrast to non-TM-proteins. There are also differences between TM-kernels of bacterial and eukaryotic proteins. We have constructed amino acid substitution matrices for bacterial TM-kernels, named the BATMAS (BActerial Transmembrane MAtrix of Substitutions) series. In TM-kernels, polar and charged residues, as well as proline and tyrosine, are highly conserved, whereas there are more substitutions within the group of hydrophobic residues, in contrast to non-TM-proteins that have fewer, relatively more conserved, hydrophobic residues. These results demonstrate that alignment of transmembrane proteins should be based on at least two amino acid substitution matrices, one for loops (e.g., the BLOSUM series) and one for TM-segments (the BATMAS series), and the choice of the TM-matrix should be different for eukaryotic and bacterial proteins.

ATP-Binding Cassette Transporters↗

Large-scale search of SNPs for type 2 DM susceptibility genes in a Japanese population.

The etiology of type 2 diabetes (DM) is polygenic. We investigated here genes and polymorphisms that associate with DM in the Japanese population. Single-nucleotide polymorphisms (SNPs) of 398 derived from 120 candidate genes were examined for association with DM in a population-based case-control study. The study group consisted of 148 cases and 227 controls recruited from Funagata, Japan. No evident subpopulation structure was detected for the tested population. The association tests were conducted with standard allele positivity tables (chi(2) tests) between SNP genotype frequency and case-control status. The independent association of the SNPs from serum triglyceride levels and body mass index was examined by multiple logistic regression analysis. A value of P<0.01 was accepted as statistically significant. Six genes (met proto-oncogene, ATP-binding cassette transporter A1, fatty acid binding protein 2, LDL receptor defect C complementing, aldolase B, and sulfonylurea receptor) were shown to be associated with DM.

ATP-Binding Cassette Transporters↗