PubMed Health⌕ Search

PubMed · 10592270

Human immunodeficiency virus reverse transcriptase and protease sequence database.

Abstract

The HIV RT and Protease Sequence Database is an online relational database that catalogs evolutionary and drug-related human immunodeficiency virus (HIV) reverse transcriptase (RT) and protease sequence variation (http://hivdb.stanford.edu). The database contains a compilation of nearly all published HIV RT and protease sequences including International Collaboration database submissions (e.g., GenBank) and sequences published in journal articles. Sequences are linked to data about the source of the sequence sample and the antiretroviral drug treatment history of the individual from whom the isolate was obtained. The database is curated and sequences are annotated with data from >230 literature references. Users can retrieve additional data and view alignments of sequence sets meeting specific criteria (e.g., treatment history, subtype, presence of a particular mutation). A gene-specific sequence analysis program, new user-defined queries and nearly 2000 additional sequences were added in 1999.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

R W Shafer, D R Jung, B J Betts, Y Xi, M J Gonzales. 2000-01-01. Human immunodeficiency virus reverse transcriptase and protease sequence database.. https://doi.org/10.1093/nar%2F28.1.346

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Protein structure comparison using the markov transition model of evolution.

A number of automatic protein structure comparison methods have been proposed; however, their similarity score functions are often decided by the researchers' intuition and trial-and-error, and not by theoretical background. We propose a novel theory to evaluate protein structure similarity, which is based on the Markov transition model of evolution. Our similarity score between structures i and j is defined as log P(j --> i)/P(i), where P(j --> i) is the probability that structure j changes to structure i during the evolutionary process, and P(i) is the probability that structure i appears by chance. This is a reasonable definition of structure similarity, especially for finding evolutionarily related (homologous) similarity. The probability P(j --> i) is estimated by the Markov transition model, which is similar to the Dayhoff's substitution model between amino acids. To estimate the parameters of the model, homologous protein structure pairs are collected using sequence similarity, and the numbers of structure transitions within the pairs are counted. Next these numbers are transformed to a transition probability matrix of the Markov transition. Transition probabilities for longer time are obtained by multiplying the probability matrix by itself several times. In this study, we generated three types of structure similarity scores: an environment score, a residue-residue distance score, and a secondary structure elements (SSE) score. Using these scores, we developed the structure comparison program, Matras (MArkovian TRAnsition of protein Structure). It employs a hierarchical alignment algorithm, in which a rough alignment is first obtained by SSEs, and then is improved with more detailed functions. We attempted an all-versus-all comparison of the SCOP database, and evaluated its ability to recognize a superfamily relationship, which was manually assigned to be homologous in the SCOP database. A comparison with the FSSP database shows that our program can recognize more homologous similarity than FSSP. We also discuss the reliability of our method, by studying the disagreement between structural classifications by Matras and SCOP.

Databases, Factual↗

Identification of lorazepam and sildenafil as examples for the application of LC/ionspray-MS and MS-MS with mass spectra library searching in forensic toxicology.

A mass spectra (MS) library using in-source collision induced dissociation (ESI-CID) as well as a tandem-mass spectra (MS-MS) library with product ion spectra of drugs has recently been developed with a triple-quadrupole ionspray mass spectrometer [1,2]. For the ESI-CID MS library, single-quadrupole mode and for the MS-MS library triple-quadrupole mode have been used. These mass spectra libraries were applied successfully for the general-unknown screening for drugs and metabolites in serum and urine with liquid-chromatography-mass spectrometry (LC-MS) using a PE/SCIEX API 365 with a turboionspray source. As examples, the identification of lorazepam and lorazepam-glucuronide in a serum extract and the identification of sildenafil and alkyloxidated sildenafil in urine are presented here.

Databases, Factual↗