PubMed HealthSearch

Biomedical subjects

R Rogiers

Publications and source records attributed to R Rogiers.

6 recordsLinked to original sources

Complete nucleotide sequence of SV40 DNA.

The determination of the total 5,224 base-pair DNA sequence of the virus SV40 has enabled us to locate precisely the known genes on the genome. At least 15.2% of the genome is presumably not translated into polypeptides. Particular points of interest revealed by the complete sequence are the initiation of the early t and T antigens at the same position and the fact that the T antigen is coded by two non-contiguous regions of the genome; the T antigen mRNA is spliced in the coding region. In the late region the gene for the major protein VP1 overlaps those for proteins VP2 and VP3 over 122 nucleotides but is read in a different frame. The almost complete amino acid sequences of the two early proteins as well as those of the late proteins have been deduced from the nucleotide sequence. The mRNAs for the latter three proteins are presumably spliced out of a common primary RNA transcript. The use of degenerate codons is decidedly non-random, but is similar for the early and late regions. Codons of the type NUC, NCG and CGN are absent or very rare.

Antigens, Neoplasm

Nucleotide sequence of the simian virus 40 Hind-K restriction fragment.

The restriction fragment Hind-K represents 4.2% of the genome of Simian virus 40 (SV40) and is located near the middle of the late region. Its nucleotide sequence is reported here. It was mainly established by analysis of transcription products, synthesized by means of Escherichia coli RNA polymerase and nucleoside triphosphates, one of which was (alpha-32P)-labeled. Strand assignment was possible by hybridization of asymmetric, labeled transcripts of total SV40 DNA to filter-bound Hind-K fragment. Further information and unambiguous confirmation of the sequence was obtained by the use of direct DNA-sequencing methods. For this purpose the fragment was labeled at the 5' ends by means of polynucleotide kinase and [gamma-32P]ATP and redigested with a suitable restriction enzyme. The separated products were then either partially digested with snake venom diesterase for analysis by the 'wandering spot' method or partially degraded with the base-specific reagents dimethylsulphate or hydrazine for direct sequence analysis on gel. The Hind-K sequence is 219 base pairs long. The message strand is particularly rich in adenosine (39%) and purines. The nucleotide sequence cna unambiguously be translated into an amino acid sequence and the N-terminal codon of the viral protein VP1 gene could be identified. The amino-terminal part of VP1 is rich in proline and lysine. The nucleotide sequence of Hind-K codes also for the carboxyl-terminal part of the viral protein VP2 and VP3 genes, which partly overlap the VP1 gene.

Base Sequence

Overlapping of the VP2-VP3 gene and the VP1 gene in the SV40 genome.

The nucleotide sequence of the SV40 Hind E fragment has been determined mainly by the partial chemical degradation procedure of Maxam and Gilbert (1977). The sequence of the strand with the same polarity as the late messenger RNA shows only one open reading frame for translation. Considering that VP3 corresponds to the carbosyl terminal part of VP2, and considering various evidence which indicates that the SV40 Hind E segment is part of the amino acid sequence of VP2-VP3. It continues clockwise in Hind K, where it terminates with a UAA signal. The latter is located 110 nucleotides beyond the initiation signal for the major structural protein VP1 (Fiers et al., 1975; Van de Voorde et al., 1976). Hence this small overlapping region of the genome codes for the synthesis of three different proteins in two different reading frames. The deduced amino acid sequence covers a major part of the vp3 poly peptide, and the amino acid composition is in good agreement with published values (Greenaway and Levine, 1973).

Base Sequence

The initiation region of the SV40 VP1 gene.

The sequence of 15 nucleotides located at the 5' terminus of the plus strand of the SV40 Hind K fragment has been determined as (5') A-G-C-T-T-A-T-G-A-A-G-A-T-G-G (3'). The 3' on OH terminal G of this segment is part of the G-C-C codeword for the N terminal alanine of the VP1 protein. This region therefore presumably corresponds to a ribosome binding site on the 16S late mRNA. Complementarily to the 3' OH of eucaryotic 18S ribosomal RNA and homology with the BMV coat ribosome binding site are discussed.

Bacterial Proteins