PubMed Health⌕ Search

PubMed · 11125091

The EMOTIF database.

Abstract

The EMOTIF database is a collection of more than 170 000 highly specific and sensitive protein sequence motifs representing conserved biochemical properties and biological functions. These protein motifs are derived from 7697 sequence alignments in the BLOCKS+ database (released on June 23, 2000) and all 8244 protein sequence alignments in the PRINTS database (version 27.0) using the emotif-maker algorithm developed by Nevill-Manning et al. (Nevill-Manning,C.G., Wu,T.D. and Brutlag,D.L. (1998) Proc. Natl Acad. Sci. USA, 95, 5865-5871; Nevill-Manning,C.G., Sethi,K.S., Wu,T. D. and Brutlag,D.L. (1997) ISMB-97, 5, 202-209). Since the amino acids and the groups of amino acids in these sequence motifs represent critical positions conserved in evolution, search algorithms employing the EMOTIF patterns can identify and classify more widely divergent sequences than methods based on global sequence similarity. The emotif protein pattern database is available at http://motif.stanford.edu/emotif/.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

J Y Huang, D L Brutlag. 2001-01-01. The EMOTIF database.. https://doi.org/10.1093/nar%2F29.1.202

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

GAMMA: gap-aware motif mining under incomplete labeling with applications to MHC motifs.

MOTIVATION: Sequence motif identification is crucial for understanding molecular recognition, particularly in immune responses involving peptide binding to major histocompatibility complex (MHC) Class I molecules for antigen presentation to T cells. Traditionally, MHC Class I binding motifs are assumed to be contiguous and span nine amino acids. However, structural evidence suggests that binding may involve nonadjacent residues, challenging the assumptions of existing methods. RESULTS: In this study, we propose Gap-Aware Motif Mining Algorithm (GAMMA), a probabilistic framework designed to identify noncontiguous motifs under conditions of incomplete labeling. GAMMA employs Bayesian inference with Markov chain Monte Carlo sampling to jointly estimate motif parameters, binding locations, and the relative spacing between binding positions. Through extensive simulations and real-world applications to MHC Class I peptide datasets, GAMMA outperforms existing motif discovery tools such as GLAM2 in accurately localizing binding residues and identifying the underlying motifs. Notably, our results suggest that the true number of binding residues may be eight, fewer than the commonly assumed nine. In addition, for longer peptides, the model captures increased flexibility in the central region, consistent with structural observations that peptides may bulge in the middle. AVAILABILITY AND IMPLEMENTATION: The raw data and the source codes are available on GitHub (https://github.com/RanLIUaca/GAMMAmotif).

Amino Acid Motifs↗

Signalling thresholds and negative B-cell selection in acute lymphoblastic leukaemia.

B cells are selected for an intermediate level of B-cell antigen receptor (BCR) signalling strength: attenuation below minimum (for example, non-functional BCR) or hyperactivation above maximum (for example, self-reactive BCR) thresholds of signalling strength causes negative selection. In ∼25% of cases, acute lymphoblastic leukaemia (ALL) cells carry the oncogenic BCR-ABL1 tyrosine kinase (Philadelphia chromosome positive), which mimics constitutively active pre-BCR signalling. Current therapeutic approaches are largely focused on the development of more potent tyrosine kinase inhibitors to suppress oncogenic signalling below a minimum threshold for survival. We tested the hypothesis that targeted hyperactivation--above a maximum threshold--will engage a deletional checkpoint for removal of self-reactive B cells and selectively kill ALL cells. Here we find, by testing various components of proximal pre-BCR signalling in mouse BCR-ABL1 cells, that an incremental increase of Syk tyrosine kinase activity was required and sufficient to induce cell death. Hyperactive Syk was functionally equivalent to acute activation of a self-reactive BCR on ALL cells. Despite oncogenic transformation, this basic mechanism of negative selection was still functional in ALL cells. Unlike normal pre-B cells, patient-derived ALL cells express the inhibitory receptors PECAM1, CD300A and LAIR1 at high levels. Genetic studies revealed that Pecam1, Cd300a and Lair1 are critical to calibrate oncogenic signalling strength through recruitment of the inhibitory phosphatases Ptpn6 (ref. 7) and Inpp5d (ref. 8). Using a novel small-molecule inhibitor of INPP5D (also known as SHIP1), we demonstrated that pharmacological hyperactivation of SYK and engagement of negative B-cell selection represents a promising new strategy to overcome drug resistance in human ALL.

Amino Acid Motifs↗

Purification and cDNA cloning of a novel antibacterial peptide with a cysteine-stabilized alphabeta motif from the longicorn beetle, Acalolepta luxuriosa.

An antibacterial peptide from the hemolymph of a coleopteran insect, Acalolepta luxuriosa, in the superfamily Cerambyocidea was characterized. The mature antibacterial peptide had 27 amino acid residues with a theoretical molecular weight of 3099.29 and it showed antibacterial activity against Escherichia coli and Micrococcus luteus. The deduced amino acid sequence of the peptide showed that it had a cysteine-stabilized alphabeta motif with a C...CXXXC...C...CXC consensus sequence, like insect defensins. However, the results of a multiple sequence alignment and phylogenetic analysis with CLUSTAL X indicated that this peptide is a novel peptide with a cysteine-stabilized alphabeta motif that is distant from insect defensins.

Amino Acid Motifs↗