PubMed Health⌕ Search

PubMed · 11751234

Sequence type analysis and recombinational tests (START).

Abstract

UNLABELLED: The 32-bit Windows application START is implemented using Visual Basic and C(++) and performs analyses to aid in the investigation of bacterial population structure using multilocus sequence data. These analyses include data summary, lineage assignment, and tests for recombination and selection. AVAILABILITY: START is available at http://outbreak.ceid.ox.ac.uk/software.htm. CONTACT: keith.jolley@ceid.ox.ac.uk

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

K A Jolley, E J Feil, M S Chan, M C Maiden. 2001. Sequence type analysis and recombinational tests (START).. https://doi.org/10.1093/bioinformatics%2F17.12.1230

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

The INSDC specifications-foundations for a FAIR and global INSDC.

Members of the International Nucleotide Sequence Database Collaboration (INSDC; https://www.insdc.org/) collect, exchange, and preserve comprehensive open nucleotide sequence information and provide tools for its access. The INSDC has stated its commitment to welcoming new members into the collaboration to be more representative of the global community of data and users. To reach this goal, a comprehensive definition of the INSDC data model and minimum requirements for data acceptance have been established. Here we describe the processes used to arrive upon these INSDC Specifications and lay out strategies for their continued upkeep to remain current and relevant. Database URL:  https://www.insdc.org/.

Databases, Nucleic Acid↗

Using expressed sequence tag databases to identify ovarian genes of interest.

GenBank contains 4879 expressed sequence tags (EST) derived from four non-normalized human ovarian cDNA libraries. Of these EST, 2646 are contributors to UniGene clusters and have UniGene numbers. The EST map to 1206 distinct UniGenes. A gene expression profile was established for the human ovary by identifying the abundance of each UniGene cluster and its corresponding annotation. The most highly expressed transcripts were for proteins associated with protein synthesis (ribosomal proteins, elongation factors, thymosins, etc.). However, there are also transcripts for genes of unknown function that are ovary-specific. This ovarian gene expression profile provides useful data for the design of DNA microarrays targeted at ovarian function and highlights novel sequences that warrant further investigation.

Databases, Nucleic Acid↗

Identification of Ugandan HIV type 1 variants with unique patterns of recombination in pol involving subtypes A and D.

Most HIV-1 infections in Uganda are caused by subtypes A and D. The prevalence of recombination and the sites of specific breakpoints between these subtypes have not been reported. HIV-1 pol sequences encoding protease (amino acids 1-99) and reverse transcriptase (amino acids 1-324) from 102 pregnant Ugandan women were analyzed by the Recombinant Identification Program, SimPlot, and examination of phylogenetically informative sites to identify sites of recombination between sequence segments belonging to different subtypes. Thirteen percent (13 of 102) of the pol sequences contained strong evidence of recombination between subtypes A and D. At least nine different patterns of recombination were observed. Five women infected with a recombinant virus transmitted the recombinant virus perinatally. In this population-based study, intersubtype recombinants were common. The large number of different types of pol recombinants identified suggests that recombination occurs readily in the pol region. Perinatal transmission of the recombinant viruses demonstrates their evolutionary stability.

Databases, Nucleic Acid↗