PubMed Health⌕ Search

Biomedical subjects

Jonathan L Tupy

Publications and source records attributed to Jonathan L Tupy.

3 recordsLinked to original sources

Identification of putative noncoding polyadenylated transcripts in Drosophila melanogaster.

Analysis of EST and cDNA collections from a number of metazoan species has identified genes encoding long polyadenylated transcripts that do not contain ORFs of lengths typical for protein-encoding mRNAs. Noncoding functions of such polyadenylated transcripts have been elucidated in only a few examples. The corresponding genes neither contain hallmark sequence motifs nor appear to have been conserved across phyla. Thus, it is impossible to systematically identify new members of this class of gene by using sequence homology and traditional gene-finding algorithms that depend on protein-coding potential. Consequently, even their approximate number has not been established for any metazoan genome. We curated polyadenylated transcripts with limited protein-coding capacity from intergenic regions of the Drosophila melanogaster genome. We used RT-PCR assays, hybridization to RNA blots and whole-mount embryos, and computational analyses to characterize candidate transcripts. We verify the structures and expression of 17 distinct, likely non-protein-coding polyadenylated transcripts. We show that the expression of many of these transcripts is conserved in other Drosophila species, indicating that they have important biological functions.

Animals↗

Crystal structure of Escherichia coli sigmaE with the cytoplasmic domain of its anti-sigma RseA.

The sigma factors are the key regulators of bacterial transcription. ECF (extracytoplasmic function) sigma's are the largest and most divergent group of sigma(70) family members. ECF sigma's are normally sequestered in an inactive complex by their specific anti-sigma factor, which often spans the inner membrane. Here, we determined the 2 A resolution crystal structure of the Escherichia coli ECF sigma factor sigma(E) in an inhibitory complex with the cytoplasmic domain of its anti-sigma, RseA. Despite extensive sequence variability, the two major domains of sigma(E) are virtually identical in structure to the corresponding domains of other sigma(70) family members. In combination with a model of the sigma(E) holoenzyme and biochemical data, the structure reveals that RseA functions by sterically occluding the two primary binding determinants on sigma(E) for core RNA polymerase.

Binding Sites↗

Annotation of the Drosophila melanogaster euchromatic genome: a systematic review.

BACKGROUND: The recent completion of the Drosophila melanogaster genomic sequence to high quality and the availability of a greatly expanded set of Drosophila cDNA sequences, aligning to 78% of the predicted euchromatic genes, afforded FlyBase the opportunity to significantly improve genomic annotations. We made the annotation process more rigorous by inspecting each gene visually, utilizing a comprehensive set of curation rules, requiring traceable evidence for each gene model, and comparing each predicted peptide to SWISS-PROT and TrEMBL sequences. RESULTS: Although the number of predicted protein-coding genes in Drosophila remains essentially unchanged, the revised annotation significantly improves gene models, resulting in structural changes to 85% of the transcripts and 45% of the predicted proteins. We annotated transposable elements and non-protein-coding RNAs as new features, and extended the annotation of untranslated (UTR) sequences and alternative transcripts to include more than 70% and 20% of genes, respectively. Finally, cDNA sequence provided evidence for dicistronic transcripts, neighboring genes with overlapping UTRs on the same DNA sequence strand, alternatively spliced genes that encode distinct, non-overlapping peptides, and numerous nested genes. CONCLUSIONS: Identification of so many unusual gene models not only suggests that some mechanisms for gene regulation are more prevalent than previously believed, but also underscores the complex challenges of eukaryotic gene prediction. At present, experimental data and human curation remain essential to generate high-quality genome annotations.

Animals↗