PubMed Health⌕ Search

PubMed · 11751226

Taxonomy workbench.

Abstract

UNLABELLED: At advanced stages of working with user-defined protein and gene sequence collections, it is frequently necessary to link these data to the taxonomic tree and to extract subsets in accordance with taxonomic considerations. Since no general automatic tools had been available, this was a tedious manual effort. Our taxonomy workbench allows processing of sequence sets, mapping of these sets onto the taxonomic tree, collection of taxonomic subsets from them and printing of the whole tree or some part of it. As a side effect, the system enables queries to and navigation within the taxonomy database. AVAILABILITY: An implementation of the taxonomy workbench is accessible for public use as a www-service at http://mendel.imp.univie.ac.at/taxonomy/. Software components for the command-line and for the www-version are available on request. CONTACT: Georg.Schneider@nt.imp.univie.ac.at; Frank.Eisenhaber@nt.imp.univie.ac.at SUPPLEMENTARY INFORMATION: Documentation for the taxonomy workbench can be accessed at http://mendel.imp.univie.ac.at/taxonomy/help.html.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

M Wildpaner, G Schneider, A Schleiffer, F Eisenhaber. 2001. Taxonomy workbench.. https://doi.org/10.1093/bioinformatics%2F17.12.1179

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

The INSDC specifications-foundations for a FAIR and global INSDC.

Members of the International Nucleotide Sequence Database Collaboration (INSDC; https://www.insdc.org/) collect, exchange, and preserve comprehensive open nucleotide sequence information and provide tools for its access. The INSDC has stated its commitment to welcoming new members into the collaboration to be more representative of the global community of data and users. To reach this goal, a comprehensive definition of the INSDC data model and minimum requirements for data acceptance have been established. Here we describe the processes used to arrive upon these INSDC Specifications and lay out strategies for their continued upkeep to remain current and relevant. Database URL:  https://www.insdc.org/.

Databases, Nucleic Acid↗

Using expressed sequence tag databases to identify ovarian genes of interest.

GenBank contains 4879 expressed sequence tags (EST) derived from four non-normalized human ovarian cDNA libraries. Of these EST, 2646 are contributors to UniGene clusters and have UniGene numbers. The EST map to 1206 distinct UniGenes. A gene expression profile was established for the human ovary by identifying the abundance of each UniGene cluster and its corresponding annotation. The most highly expressed transcripts were for proteins associated with protein synthesis (ribosomal proteins, elongation factors, thymosins, etc.). However, there are also transcripts for genes of unknown function that are ovary-specific. This ovarian gene expression profile provides useful data for the design of DNA microarrays targeted at ovarian function and highlights novel sequences that warrant further investigation.

Databases, Nucleic Acid↗

Identification of Ugandan HIV type 1 variants with unique patterns of recombination in pol involving subtypes A and D.

Most HIV-1 infections in Uganda are caused by subtypes A and D. The prevalence of recombination and the sites of specific breakpoints between these subtypes have not been reported. HIV-1 pol sequences encoding protease (amino acids 1-99) and reverse transcriptase (amino acids 1-324) from 102 pregnant Ugandan women were analyzed by the Recombinant Identification Program, SimPlot, and examination of phylogenetically informative sites to identify sites of recombination between sequence segments belonging to different subtypes. Thirteen percent (13 of 102) of the pol sequences contained strong evidence of recombination between subtypes A and D. At least nine different patterns of recombination were observed. Five women infected with a recombinant virus transmitted the recombinant virus perinatally. In this population-based study, intersubtype recombinants were common. The large number of different types of pol recombinants identified suggests that recombination occurs readily in the pol region. Perinatal transmission of the recombinant viruses demonstrates their evolutionary stability.

Databases, Nucleic Acid↗