PubMed Health⌕ Search

PubMed · 10592220

The IDB and IEDB: intron sequence and evolution databases.

Abstract

A non-redundant database of nuclear, protein-encoding, genomic DNA sequences highlighting nuclear pre-mRNA introns was constructed using information contained in the SWISS-PROT and GenBank sequence databases. This Intron DataBase (IDB) contains information about (i) introns (including nucleotide sequence, location, phase, length, GC content and consensus-sequence rule violations), (ii) exons (including nucleo-tide sequence, length and GC content), (iii) protein coding regions (including amino acid sequence and length), and (iv) descriptive information about the source gene and organism (including gene designations and species taxonomy). The Intron Evolution DataBase (IEDB) provides a statistical analysis of the exon and intron sequences catalogued in IDB as well as data concerning intron penetration (relative number of coding regions with introns), density (number of introns per kb of total coding sequence DNA), distribution, and consensus sequences for each species present in IDB. This supplement is provided to furnish insights into the phylogenetic distribution and evolution of introns. Both databases are extensively cross-referenced to the SWISS-PROT and GenBank databases. IDB currently contains information on over 63 000 genes and 154 000 introns; IEDB summarizes information on over 2800 species. IDB and IEDB will be updated twice a year and are available via the internet (http://nutmeg.bio.indiana. edu/intron/index.html ).

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

N J Schisler, J D Palmer. 2000-01-01. The IDB and IEDB: intron sequence and evolution databases.. https://doi.org/10.1093/nar%2F28.1.181

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

miR-503-3p promotes epithelial-mesenchymal transition in breast cancer by directly targeting SMAD2 and E-cadherin.

Although progress in clinical and basic research has significantly increased our understanding of breast cancer, little is known about the molecular mechanism underlying breast cancer metastasis. Identification of effective therapeutic targets to prevent breast cancer metastasis is urgently needed. The function of miR-503-3p has been investigated in other cancers, but its role in breast cancer remains undefined. Here, we found that miR-503-3p was overexpressed in breast cancer tissue and plasma compared with adjacent normal breast tissue and with plasma from healthy individuals. Moreover, we identified miR-503-3p to be an oncogene of breast cancer cell proliferation, migration and invasion. Upregulation of miR-503-3p in breast cancer cells inhibited expression of epithelial-mesenchymal transition (EMT)-related protein SMAD2 and the epithelial marker protein E-cadherin by directly binding to their mRNA 3' untranslated region, whereas increased expression of mesenchymal marker proteins, including vimentin and N-cadherin. Taken together, our findings support a critical role for miR-503-3p in induction of breast cancer EMT and suggest that plasma miR-503-3p may be a useful diagnostic biomarker for breast cancer.

Base Sequence↗

Detection and identification of Escherichia coli, Shigella, and Salmonella by microarrays using the gyrB gene.

Commonly, 16S ribosome RNA (16S rRNA) sequence analysis has been used for identifying enteric bacteria. However, it may not always be applicable for distinguishing closely related bacteria. Therefore, we selected gyrB genes that encode the subunit B protein of DNA gyrase (a topoisomerase type II protein) as target genes. The molecular evolution rate of gyrB genes is higher than that of 16S rRNA, and gyrB genes are distributed universally among bacterial species. Microarray technology includes the methods of arraying cDNA or oligonucleotides on substrates such as glass slides while acquiring a lot of information simultaneously. Thus, it is possible to identify the enteric bacteria easily using microarray technology. We devised a simple method of rapidly identifying bacterial species through the combined use of gyrB genes and microarrays. Closely related bacteria were not identified at the species level using 16S rRNA sequence analysis, whereas they were identified at the species level based on the reaction patterns of oligonucleotides on our microarrays using gyrB genes.

Base Sequence↗

Ribonuclease III-mediated processing of specific Neisseria meningitidis mRNAs.

Approx. 2% of the Neisseria meningitidis genome consists of small DNA insertion sequences known as Correia or nemis elements, which feature TIRs (terminal inverted repeats) of 26-27 bp in length. Elements interspersed with coding regions are co-transcribed with flanking genes into mRNAs, processed at double-stranded RNA structures formed by TIRs. N. meningitidis RNase III (endoribonuclease III) is sufficient to process nemis+ RNAs. RNA hairpins formed by nemis with the same termini (26/26 and 27/27 repeats) are cleaved. By contrast, bulged hairpins formed by 26/27 repeats inhibit cleavage, both in vitro and in vivo. In electrophoretic mobility shift assays, all hairpin types formed similar retarded complexes upon incubation with RNase III. The levels of corresponding nemis+ and nemis- mRNAs, and the relative stabilities of RNA segments processed from nemis+ transcripts in vitro, may both vary significantly.

Base Sequence↗