PubMed Health⌕ Search

Biomedical subjects

Hanah Margalit

Publications and source records attributed to Hanah Margalit.

5 recordsLinked to original sources

How reliable are experimental protein-protein interaction data?

Data of protein-protein interactions provide valuable insight into the molecular networks underlying a living cell. However, their accuracy is often questioned, calling for a rigorous assessment of their reliability. The computation offered here provides an intelligible mean to assess directly the rate of true positives in a data set of experimentally determined interacting protein pairs. We show that the reliability of high-throughput yeast two-hybrid assays is about 50%, and that the size of the yeast interactome is estimated to be 10,000-16,600 interactions.

Protein Binding↗

Conserved sequence elements associated with exon skipping.

One of the major forms of alternative splicing, which generates multiple mRNA isoforms differing in the precise combinations of their exon sequences, is exon skipping. While in constitutive splicing all exons are included, in the skipped pattern(s) one or more exons are skipped. The regulation of this process is still not well understood; so far, cis- regulatory elements (such as exonic splicing enhancers) were identified in individual cases. We therefore set to investigate the possibility that exon skipping is controlled by sequences in the adjacent introns. We employed a computer analysis on 54 sequences documented as undergoing exon skipping, and identified two motifs both in the upstream and downstream introns of the skipped exons. One motif is highly enriched in pyrimidines (mostly C residues), and the other motif is highly enriched in purines (mostly G residues). The two motifs differ from the known cis-elements present at the 5' and 3' splice site. Interestingly, the two motifs are complementary, and their relative positional order is conserved in the flanking introns. These suggest that base pairing interactions can underlie a mechanism that involves secondary structure to regulate exon skipping. Remarkably, the two motifs are conserved in mouse orthologous genes that undergo exon skipping.

Alternative Splicing↗

A survey of small RNA-encoding genes in Escherichia coli.

Small RNA (sRNA) molecules have gained much interest lately, as recent genome-wide studies have shown that they are widespread in a variety of organisms. The relatively small family of 10 known sRNA-encoding genes in Escherichia coli has been significantly expanded during the past two years with the discovery of 45 novel genes. Most of these genes are still uncharacterized and their cellular roles are unknown. In this survey we examined the sequence and genomic features of the 55 currently known sRNA-encoding genes in E.coli, attempting to identify their common characteristics. Such characterization is important for both expanding our understanding of this unique gene family and for improving the methods to predict and identify sRNA-encoding genes based on genomic information.

Base Composition↗

PeCoP: automatic determination of persistently conserved positions in protein families.

UNLABELLED: PeCoP is a WWW-based service which accepts a protein sequence, and reports positions that are conserved in close and distant sequence family members. The collation of family members is performed using iterative PSI-BLAST runs. Examining positional conservation in close and distant family members enables a better selection of positions that may play a role in determining the protein's structure and function. AVAILABILITY: http://bioinformatics.org/pecop CONTACT: idoerg@burnham.org SUPPLEMENTARY INFORMATION: http://bioinformatics.org/pecop/about_pecop

Amino Acid Sequence↗

Persistently conserved positions in structurally similar, sequence dissimilar proteins: roles in preserving protein fold and function.

Many protein pairs that share the same fold do not have any detectable sequence similarity, providing a valuable source of information for studying sequence-structure relationship. In this study, we use a stringent data set of structurally similar, sequence-dissimilar protein pairs to characterize residues that may play a role in the determination of protein structure and/or function. For each protein in the database, we identify amino-acid positions that show residue conservation within both close and distant family members. These positions are termed "persistently conserved". We then proceed to determine the "mutually" persistently conserved (MPC) positions: those structurally aligned positions in a protein pair that are persistently conserved in both pair mates. Because of their intra- and interfamily conservation, these positions are good candidates for determining protein fold and function. We find that 45% of the persistently conserved positions are mutually conserved. A significant fraction of them are located in critical positions for secondary structure determination, they are mostly buried, and many of them form spatial clusters within their protein structures. A substitution matrix based on the subset of MPC positions shows two distinct characteristics: (i) it is different from other available matrices, even those that are derived from structural alignments; (ii) its relative entropy is high, emphasizing the special residue restrictions imposed on these positions. Such a substitution matrix should be valuable for protein design experiments.

Amino Acid Motifs↗