PubMed HealthSearch

Biomedical subjects

P Bucher

Publications and source records attributed to P Bucher.

At least 19 recordsLinked to original sources

The PROSITE database, its status in 1995.

The PROSITE database consists of biologically significant patterns and profiles formulated in such a way that with appropriate computational tools it can help to determine to which known family of proteins (if any) a new sequence belongs or which known domain(s) it contains.

Computer Communication Networks

A flexible motif search technique based on generalized profiles.

A flexible motif search technique is presented which has two major components: (1) a generalized profile syntax serving as a motif definition language; and (2) a motif search method specifically adapted to the problem of finding multiple instances of a motif in the same sequence. The new profile structure, which is the core of the generalized profile syntax, combines the functions of a variety of motif descriptors implemented in other methods, including regular expression-like patterns, weight matrices, previously used profiles, and certain types of hidden Markov models (HMMs). The relationship between generalized profiles and other biomolecular motif descriptors is analyzed in detail, with special attention to HMMs. Generalized profiles are shown to be equivalent to a particular class of HMMs, and conversion procedures in both directions are given. The conversion procedures provide an interpretation for local alignment in the framework of stochastic models, allowing for clear, simple significance tests. A mathematical statement of the motif search problem defines the new method exactly without linking it to a specific algorithmic solution. Part of the definition includes a new definition of disjointness of alignments.

Algorithms

Single-cell PCR analysis of TCR repertoires selected by antigen in vivo: a high magnitude CD8 response is comprised of very few clones.

Taking advantage of a potent MHC class I-restricted response that allows the identification of antigen-selected CD8 T cells directly ex vivo, we characterized the antigen-specific T cell repertoires that develop in individual mice by single-cell PCR analysis. Each of the immune mice displayed distinct yet structurally similar TCR repertoires. The overall repertoire size was estimated to be in the range of 15-20 for most mice. No major differences were observed between primary and secondary responses. Moreover, for a hyperimmunized mouse the antigen-specific TCR repertoire expressed 8 months after the initial immunization was very similar to that found at the peak of the primary response. Our results demonstrate that a high magnitude immune response may be composed of very few clones, and that at least in the system analyzed, the memory response largely reflects the repertoire selected by the peak of the primary response.

Amino Acid Sequence

Memory TCR repertoires analyzed long-term reflect those selected during the primary response.

Normal T cell repertoire selection and evolution in antigen-specific responses is poorly understood. We have recently described an MHC class I-restricted response characterized by an overwhelming expansion of CD8 cells expressing a Vbeta10 TCR, thus allowing the identification of antigen-selected cells directly ex vivo. Our present strategy to follow the overall TCR repertoire selection was to monitor the expression of a particular TCR alpha chain (Valpha8) on antigen-selected Vbeta10(+) cells by four-color flow cytometry. We demonstrate that while there is substantial variation among the responder mice in Valpha8 usage, the repertoires of individual animals remain relatively stable over long periods of time (>1 year), with or without repeated antigenic challenge. Thus if any evolution of this response occurs upon re-exposure to antigen, it would appear not to skew the TCR repertoire established during the primary response.

Animals

A consensus motif in the RFX DNA binding domain and binding domain mutants with altered specificity.

The RFX DNA binding domain is a novel motif that has been conserved in a growing number of dimeric DNA-binding proteins, having diverse regulatory functions, in eukaryotic organisms ranging from yeasts to humans. To characterize this novel motif, we have performed a detailed dissection of the site-specific DNA binding activity of RFX1, a prototypical member of the RFX family. First, we have performed a site selection procedure to define the consensus binding site of RFX1. Second, we have developed a new mutagenesis-selection procedure to derive a precise consensus motif, and to test the accuracy of a secondary structure prediction, for the RFX domain. Third, a modification of this procedure has allowed us to isolate altered-specificity RFX1 mutants. These results should facilitate the identification both of additional candidate genes controlled by RFX1 and of new members of the RFX family. Moreover, the altered-specificity RFX1 mutants represent valuable tools that will permit the function of RFX1 to be analyzed in vivo without interference from the ubiquitously expressed endogenous protein. Finally, the simplicity, efficiency, and versatility of the selection procedure we have developed make it of general value for the determination of consensus motifs, and for the isolation of mutants exhibiting altered functional properties, for large protein domains involved in protein-DNA as well as protein-protein interactions.

Amino Acid Sequence

[Reproductive medicine in transition: new developments in embryo transfer].

Successful application of embryo transfer (ET) has become common practice in cattle, horses, sheep, goats and a variety of other species held in captivity. Yet in cattle only has the technique been established commercially. In 1994 more than 100,000 bovine embryos have been transferred in European countries. Important progress in transvaginal ovum pick up (OPU), in vitro production (IVP) and cryopreservation have further improved the applicability of ET. Direct transfer simplifies the procedure considerably allowing individual transfers and eliminating the need of synchronizing recipients. In Switzerland the organization 'Veterinary Society for Embryo Transfer' (TIGET) has been founded in 1995 to support practitioners performing embryo transfer.

Animals

A sequence similarity search algorithm based on a probabilistic interpretation of an alignment scoring system.

We present a probabilistic interpretation of local sequence alignment methods where the alignment scoring system (ASS) plays the role of a stochastic process defining a probability distribution over all sequence pairs. An explicit algorithms is given to compute the probability of two sequences given and ASS. Based on this definition, a modified version of the Smith-Waterman local similarity search algorithm has been devised, which assesses sequence relationships by log likelihood ratios. When tested on classical examples such as globins or G-protein-coupled receptors, the new method proved to be up to an order of magnitude more sensitive than the native Smith-Waterman algorithm.

Algorithms

Mouse interleukin-2 receptor alpha gene expression. Interleukin-1 and interleukin-2 control transcription via distinct cis-acting elements.

We have shown that interleukin-1 (IL-1) and IL-2 control IL-2 receptor alpha (IL-2R alpha) gene transcription in CD4-CD8- murine T lymphocyte precursors. Here we map the cis-acting elements that mediate interleukin responsiveness of the mouse IL-2R alpha gene using a thymic lymphoma-derived hybridoma (PC60). The transcriptional response of the IL-2R alpha gene to stimulation by IL-1 + IL-2 is biphasic. IL-1 induces a rapid, protein synthesis-independent appearance of IL-2R alpha mRNA that is blocked by inhibitors of NF-kappa B activation. It also primes cells to become IL-2 responsive and thereby prepares the second phase, in which IL-2 induces a 100-fold further increase in IL-2R alpha transcripts. Transient transfection experiments show that several elements in the promoter-proximal region of the IL-2R alpha gene contribute to IL-1 responsiveness, most importantly an NF-kappa B site conserved in the human and mouse gene. IL-2 responsiveness, on the other hand, depends on a 78-nucleotide segment 1.3 kilobases upstream of the major transcription start site. This segment functions as an IL-2-inducible enhancer and lies within a region that becomes DNase I hypersensitive in normal T cells in which IL-2R alpha expression has been induced. IL-2 responsiveness requires three distinct elements within the enhancer. Two of these are potential binding sites for STAT proteins.

Animals

The rsp5-domain is shared by proteins of diverse functions.

A novel, unusually small, and highly conserved domain of modular intracellular proteins is described. The domain was first recognized as three repeats in the yeast rsp5 gene product and named thereafter. The rsp5 protein is thought to interact with nuclear proteins but also contains a C2 domain typical for cytoplasmic proteins. Further analyses revealed several additional occurrences of this domain in diverse protein classes, including cytoplasmic signal transduction proteins, gene products interacting with the transcription machinery, structural proteins like dystrophin, and a putative RNA helicase.

Amino Acid Sequence

Improving the sensitivity of the sequence profile method.

The sequence profile method (Gribskov M, McLachlan AD, Eisenberg D, 1987, Proc Natl Acad Sci USA 84:4355-4358) is a powerful tool to detect distant relationships between amino acid sequences. A profile is a table of position-specific scores and gap penalties, providing a generalized description of a protein motif, which can be used for sequence alignments and database searches instead of an individual sequence. A sequence profile is derived from a multiple sequence alignment. We have found 2 ways to improve the sensitivity of sequence profiles: (1) Sequence weights: Usage of individual weights for each sequence avoids bias toward closely related sequences. These weights are automatically assigned based on the distance of the sequences using a published procedure (Sibbald PR, Argos P, 1990, J Mol Biol 216:813-818). (2) Amino acid substitution table: In addition to the alignment, the construction of a profile also needs an amino acid substitution table. We have found that in some cases a new table, the BLOSUM45 table (Henikoff S, Henikoff JG, 1992, Proc Natl Acad Sci USA 89:10915-10919), is more sensitive than the original Dayhoff table or the modified Dayhoff table used in the current implementation. Profiles derived by the improved method are more sensitive and selective in a number of cases where previous methods have failed to completely separate true members from false positives.

Amino Acid Sequence

A generalized profile syntax for biomolecular sequence motifs and its function in automatic sequence interpretation.

A general syntax for expressing biomolecular sequence motifs is described, which will be used in future releases of the PROSITE data bank and in a similar collection of nucleic acid sequence motifs currently under development. The central part of the syntax is a regular structure which can be viewed as a generalization of the profiles introduced by Gribskov and coworkers. Accessory features implement specific motif search strategies and provide information helpful for the interpretation of predicted matches. Two contrasting examples, representing E. coli promoters and SH3 domains respectively, are shown to demonstrate the versatility of the syntax, and its compatibility with diverse motif search methods. It is argued, that a comprehensive machine-readable motif collection based on the new syntax, in conjunction with a standard search program, can serve as a general-purpose sequence interpretation and function prediction tool.

Animals

PROSITE: recent developments.

PROSITE is a compilation of sites and patterns found in protein sequences; it can be used as a method of determining the function of uncharacterized proteins translated from genomic or cDNA sequences.

Amino Acid Sequence

Correlation analysis of amino acid usage in protein classes.

We present a comparative study of residue usage correlations of various organism protein sets of diverse phylogenetic species and of open reading frames of several large human viral genomes. Our correlation analysis reveals three major tendencies: (i) charge compensation reflected by the high correlation of basic with acidic residues; (ii) the positive correlations of functionally and structurally similar amino acids including many pairs of hydrophobic amino acids, all pairs of aromatic amino acids, the anionic pair (glutamate and aspartate), but not the cationic pair (lysine and arginine), moderately the hydroxyl pair (serine and threonine), the small amino acids (glycine and alanine), and many (but not all) of those having high values in the Dayhoff substitutability matrix (characteristics such as amino acid polarity or codon usage agreement, except for the wobble position, do not necessarily imply significant positive correlations); (iii) a widespread negative correlation of the aggregate strong codon group amino acids (Ala, Gly, Pro) versus the weak codon group amino acids (Lys, Ile, Tyr, Asn, Phe). Discussion and speculations relate amino acid usage correlations to protein function/structure, cellular localization, proximity in amino acid biosynthetic pathways, amino acid relative abundances, tRNA and aminoacyl synthetase availabilities, and evolutionary processes.

Amino Acids

Methods and algorithms for statistical analysis of protein sequences.

We describe several protein sequence statistics designed to evaluate distinctive attributes of residue content and arrangement in primary structure. Considered are global compositional biases, local clustering of different residue types (e.g., charged residues, hydrophobic residues, Ser/Thr), long runs of charged or uncharged residues, periodic patterns, counts and distribution of homooligopeptides, and unusual spacings between particular residue types. The computer program SAPS (statistical analysis of protein sequences) calculates all the statistics for any individual protein sequence input and is available for the UNIX environment through electronic mail on request to V.B. (volker/genomic@stanford.edu).

Algorithms

Significant similarity and dissimilarity in homologous proteins.

Common practice emphasizes significant sequence similarities between different members of protein families. These similarities presumably reflect on evolutionary conservation of structurally and functionally essential residues. The nonconserved regions, on the other hand, may be either selectively neutral or differentiated. We propose several distributional sequence statistics (e.g., clustering of charged residues, compositional biases, and repetitive patterns) as indicators of differentiation events. These ideas are illustrated with various examples, including comparisons among G protein-coupled receptors, herpesvirus proteins, and GTPase-activating proteins.

GTP-Binding Proteins