PubMed Health⌕ Search

Biomedical subjects

M A Rajandream

Publications and source records attributed to M A Rajandream.

17 recordsLinked to original sources

Analysis of 41 kb of the DNA sequence from the right arm of chromosome II of Schizosaccharomyces pombe.

We report the complete sequence of cosmid c18A7 (41 046 bp insert), located on the right arm of chromosome II of the Schizosaccharomyces pombe genome. The sequence, which partially overlaps with cosmids SPBC4F6 and SPBC336, contains 16 open reading frames (ORFs) capable of coding for proteins of at least 100 amino acid residues in length (one partial) and one small nucleolar RNA (snoRNA). Four known genes were found: swi10 (encoding a mating-type switching protein also involved in nucleotide excision repair); dim1 (encoding a dimethyladenosine transferase); arf1 (encoding ADP-ribosylation factor 1); and pol3 (cdc6) the partial fragment, encoding the 125 kDa catalytic subunit of the DNA polymerase type B. Six ORFs similar to known proteins were found. They include a transporter of the major facilitator superfamily class, a vacuolar sorting protein, an asparagine synthase, a nuclear protein, a reticulum oxidoreductin and a heat shock protein. Each protein product of the other six ORFs has conserved domains and can be assigned a molecular, but not a biological, function. The sequence has been submitted to the EMBL database under Accession No. AL080287.

Amino Acid Sequence↗

A superfamily of variant genes encoded in the subtelomeric region of Plasmodium vivax.

The malarial parasite Plasmodium vivax causes disease in humans, including chronic infections and recurrent relapses, but the course of infection is rarely fatal, unlike that caused by Plasmodium falciparum. To investigate differences in pathogenicity between P. vivax and P. falciparum, we have compared the subtelomeric domains in the DNA of these parasites. In P. falciparum, subtelomeric domains are conserved and contain ordered arrays of members of multigene families, such as var, rif and stevor, encoding virulence determinants of cytoadhesion and antigenic variation. Here we identify, through the analysis of a continuous 155,711-base-pair sequence of a P. vivax chromosome end, a multigene family called vir, which is specific to P. vivax. The vir genes are present at about 600-1,000 copies per haploid genome and encode proteins that are immunovariant in natural infections, indicating that they may have a functional role in establishing chronic infection through antigenic variation.

Adult↗

Subtelomeric sequence from the right arm of Schizosaccharomyces pombe chromosome I contains seven permease genes.

The sequence has been determined of 80 888 bp of contiguous subtelomeric DNA, including the isp5 gene, from the right arm of chromosome I of Schizosaccharomyces pombe; 27 open reading frames (ORFs) longer than 100 codons are present, giving a density of one gene per 3.0 kb. Seven of the predicted proteins are members of the major facilitator superfamily (MFS) of transport proteins, including four amino acid permease homologues, bringing this family of amino acid permease sequences to 17 in Sz. pombe, and a phylogenetic analysis is presented. Also encoded is an allantoate permease homologue, a sulphate permease homologue and a probable urea active transporter. Predicted non-membrane proteins include a 1-aminocyclopropane-1-carboxylate deaminase (ACC deaminase), a class III aminotransferase, serine acetyltransferase, protein-L-isoaspartate O-methyltransferase, alpha-glucosidase, alpha-galactosidase, esterase/lipase, oxidoreductase of the short-chain dehydrogenase/reductase (SDR) family, aldehyde dehydrogenase, formamidase, amidase, flavohaemoprotein, a putative translation initiation inhibitor and a protein with similarity to a filamentous fungal conidiation-specific protein. The remaining six ORFs are likely to encode proteins, either because they have sequence similarity with hypothetical proteins or because they are known to be transcribed. Introns are scarce in the sequenced region: only three ORFs contain introns, with only one having multiple introns. The sequenced region also contains a single Tf1 transposon long terminal repeat (LTR). The sequence is derived from cosmid clones c869, c922 and c1039 and has been submitted to the EMBL database under entries SPAC869 (Accession No. AL132779), SPAC922 (AL133522) and SPAC1039 (AL133521).

Chromosomes, Fungal↗

Massive gene decay in the leprosy bacillus.

Leprosy, a chronic human neurological disease, results from infection with the obligate intracellular pathogen Mycobacterium leprae, a close relative of the tubercle bacillus. Mycobacterium leprae has the longest doubling time of all known bacteria and has thwarted every effort at culture in the laboratory. Comparing the 3.27-megabase (Mb) genome sequence of an armadillo-derived Indian isolate of the leprosy bacillus with that of Mycobacterium tuberculosis (4.41 Mb) provides clear explanations for these properties and reveals an extreme case of reductive evolution. Less than half of the genome contains functional genes but pseudogenes, with intact counterparts in M. tuberculosis, abound. Genome downsizing and the current mosaic arrangement appear to have resulted from extensive recombination events between dispersed repetitive sequences. Gene deletion and decay have eliminated many important metabolic activities including siderophore production, part of the oxidative and most of the microaerophilic and anaerobic respiratory chains, and numerous catabolic systems and their regulatory circuits.

Animals↗

Secondary DNA structure analysis of the coding strand switch regions of five Leishmania major Friedlin chromosomes.

As part of the EULEISH international genome project, a region of 74,674 nucleotides from chromosome 21 of Leishmania major Friedlin was subcloned and sequenced; and 31 new coding sequences were predicted. Of particular interest was a unique coding strand switching region covering 1.6 kb of DNA; and this was subjected to further investigation. Bioinformatic analysis of this region revealed an unusually high AT composition, a lack of putative hairpins and a strong curvature of the DNA in agreement with the structural characteristics of similar regions of other Leishmania chromosomes. These observations and a comparison with the secondary DNA structure of four other Leishmania chromosomes and chromosomes of different organisms could suggest a functional role of this region in transcription and mitotic division.

Animals↗

Prevalence of small inversions in yeast gene order evolution.

Gene order evolution in two eukaryotes was studied by comparing the Saccharomyces cerevisiae genome sequence to extensive new data from whole-genome shotgun and cosmid sequencing of Candida albicans. Gene order is substantially different between these two yeasts, with only 9% of gene pairs that are adjacent in one species being conserved as adjacent in the other. Inversion of small segments of DNA, less than 10 genes long, has been a major cause of rearrangement, which means that even where a pair of genes has been conserved as adjacent, the transcriptional orientations of the two genes relative to one another are often different. We estimate that about 1,100 single-gene inversions have occurred since the divergence between these species. Other genes that are adjacent in one species are in the same neighborhood in the other, but their precise arrangement has been disrupted, probably by multiple successive multigene inversions. We estimate that gene adjacencies have been broken as frequently by local rearrangements as by chromosomal translocations or long-distance transpositions. A bias toward small inversions has been suggested by other studies on animals and plants and may be general among eukaryotes.

Candida albicans↗

Complete DNA sequence of a serogroup A strain of Neisseria meningitidis Z2491.

Neisseria meningitidis causes bacterial meningitis and is therefore responsible for considerable morbidity and mortality in both the developed and the developing world. Meningococci are opportunistic pathogens that colonize the nasopharynges and oropharynges of asymptomatic carriers. For reasons that are still mostly unknown, they occasionally gain access to the blood, and subsequently to the cerebrospinal fluid, to cause septicaemia and meningitis. N. meningitidis strains are divided into a number of serogroups on the basis of the immunochemistry of their capsular polysaccharides; serogroup A strains are responsible for major epidemics and pandemics of meningococcal disease, and therefore most of the morbidity and mortality associated with this disease. Here we have determined the complete genome sequence of a serogroup A strain of Neisseria meningitidis, Z2491. The sequence is 2,184,406 base pairs in length, with an overall G+C content of 51.8%, and contains 2,121 predicted coding sequences. The most notable feature of the genome is the presence of many hundreds of repetitive elements, ranging from short repeats, positioned either singly or in large multiple arrays, to insertion sequences and gene duplications of one kilobase or more. Many of these repeats appear to be involved in genome fluidity and antigenic variation in this important human pathogen.

Antigenic Variation↗

The genome sequence of the food-borne pathogen Campylobacter jejuni reveals hypervariable sequences.

Campylobacter jejuni, from the delta-epsilon group of proteobacteria, is a microaerophilic, Gram-negative, flagellate, spiral bacterium-properties it shares with the related gastric pathogen Helicobacter pylori. It is the leading cause of bacterial food-borne diarrhoeal disease throughout the world. In addition, infection with C. jejuni is the most frequent antecedent to a form of neuromuscular paralysis known as Guillain-Barré syndrome. Here we report the genome sequence of C. jejuni NCTC11168. C. jejuni has a circular chromosome of 1,641,481 base pairs (30.6% G+C) which is predicted to encode 1,654 proteins and 54 stable RNA species. The genome is unusual in that there are virtually no insertion sequences or phage-associated sequences and very few repeat sequences. One of the most striking findings in the genome was the presence of hypervariable sequences. These short homopolymeric runs of nucleotides were commonly found in genes encoding the biosynthesis or modification of surface structures, or in closely linked genes of unknown function. The apparently high rate of variation of these homopolymeric tracts may be important in the survival strategy of C. jejuni.

Amino Acid Sequence↗

Analysis of 114 kb of DNA sequence from fission yeast chromosome 2 immediately centromere-distal to his5.

One hundred and fourteen kilobase pairs (kb) of contiguous genomic sequence have been determined immediately distal to the his5 genetic marker located about 0.9 Mb from the centromere on the long arm of Schizosaccharomyces pombe chromosome 2. The sequence is contained in overlapping cosmid clones c16H5, c12D12, c24C6 and c19G7, of which 20 kb are identical to previously reported sequence from clone c21H7. The remaining 93 781 bp of sequence contains 10 known genes (cdc14, cdm1, cps1, gpa1, msh2, pck2, rip1, rps30-2, sad1 and ubl1), 32 open reading frames (ORFs) capable of coding for proteins of at least 100 amino acid residues in length, one 5S rRNA gene, one tRNA(Pro) gene, one lone Tf1-type long terminal repeat (LTR) and one lone Tf2-type LTR. There is a density of one protein-coding gene per 2.2 kb and 22 of the 42 ORFs (52%) incorporate one or more introns. Twenty-one of the novel ORFs show sequence similarities which suggest functions of their products, including a cyclin C, a MADS box transcription factor, mad2-like protein, telomere binding protein, topoisomerase II-associated protein, ATP-dependent DEAH box RNA helicase, G10 protein, ubiquitin-activating e1-like enzyme, nucleoporin, prolyl-tRNA synthetase, peptidylprolyl isomerase, delta-1-pyrroline-5-carboxylate dehydrogenase, protein transport protein, coatomer epsilon, TCP-1 chaperonin, beta-subunit of 6-phosphofructokinase, aminodeoxychorismate lyase, a phosphate transport protein and a thioredoxin.

Base Sequence↗

Sequence analysis of two cosmids from Schizosaccharomyces pombe chromosome III.

We report the complete sequence of two cosmids, SPCC895 (38457 bp insert, EMBL Accession No. AL035247) and SPCC1322 (42068 bp insert, EMBL Accession No. AL035259), localized on chromosome III of the Schizosaccharomyces pombe genome. Fourteen Coding DNA sequences (CDSs) were identified in SPCC895 and 17 in SPCC1322. Two known genes were found in each cosmid: map2 and gms1 on SPCC895, encoding the mating type P-factor precursor and an UDP-galactose transporter, respectively, and bub1 and ade6 in SPCC1322, encoding a protein kinase and a phosphoribosylaminoimidazole carboxylase, respectively. The fission yeast K RNA gene has been localized to SPCC895. Three ribosomal proteins have been predicted among these two cosmids. Nine CDSs similar to known proteins were found on SPCC895, and seven on SPCC1322. They include putative genes for an uridylate kinase, a proteasome catalytic component, an ion transporter, a checkpoint protein, a translation initiation protein, a SNARE complex protein, a protein involved in cytoskeletal organization, a spindle pole body-associating protein, pre-mRNA splicing factor RNA helicase, a 3'-5' exonuclease for RNA 3' ss-tail, an UTP-glucose-1-phosphate uridylyltransferase, a leukotriene A(4) hydrolase, a member of the RanBP7-importin beta-Cse1p superfamily, a Ca(++)-calmodulin-dependent serine/threonine protein kinase and a prohibitin antiproliferative protein. One CDS is predicted to be an integral membrane protein. One CDS from SPCC895 is similar to a CDS of unknown function from Saccharomyces cerevisiae and three from SPCC1322 are similar to CDSs of unknown function from Candida albicans, S. cerevisiae and Sz. pombe, respectively. Finally, one CDS of SPCC895 and three of SPCC1322 correspond to orphan genes.

Amino Acid Sequence↗

Artemis: sequence visualization and annotation.

SUMMARY: Artemis is a DNA sequence visualization and annotation tool that allows the results of any analysis or sets of analyses to be viewed in the context of the sequence and its six-frame translation. Artemis is especially useful in analysing the compact genomes of bacteria, archaea and lower eukaryotes, and will cope with sequences of any size from small genes to whole genomes. It is implemented in Java, and can be run on any suitable platform. Sequences and annotation can be read and written directly in EMBL, GenBank and GFF format. AVAILABITLTY: Artemis is available under the GNU General Public License from http://www.sanger.ac.uk/Software/Artemis

Databases, Factual↗

Sequence and analysis of chromosome 4 of the plant Arabidopsis thaliana.

The higher plant Arabidopsis thaliana (Arabidopsis) is an important model for identifying plant genes and determining their function. To assist biological investigations and to define chromosome structure, a coordinated effort to sequence the Arabidopsis genome was initiated in late 1996. Here we report one of the first milestones of this project, the sequence of chromosome 4. Analysis of 17.38 megabases of unique sequence, representing about 17% of the genome, reveals 3,744 protein coding genes, 81 transfer RNAs and numerous repeat elements. Heterochromatic regions surrounding the putative centromere, which has not yet been completely sequenced, are characterized by an increased frequency of a variety of repeats, new repeats, reduced recombination, lowered gene density and lowered gene expression. Roughly 60% of the predicted protein-coding genes have been functionally characterized on the basis of their homology to known genes. Many genes encode predicted proteins that are homologous to human and Caenorhabditis elegans proteins.

Animals↗

The complete nucleotide sequence of chromosome 3 of Plasmodium falciparum.

Analysis of Plasmodium falciparum chromosome 3, and comparison with chromosome 2, highlights novel features of chromosome organization and gene structure. The sub-telomeric regions of chromosome 3 show a conserved order of features, including repetitive DNA sequences, members of multigene families involved in pathogenesis and antigenic variation, a number of conserved pseudogenes, and several genes of unknown function. A putative centromere has been identified that has a core region of about 2 kilobases with an extremely high (adenine + thymidine) composition and arrays of tandem repeats. We have predicted 215 protein-coding genes and two transfer RNA genes in the 1,060,106-base-pair chromosome sequence. The predicted protein-coding genes can be divided into three main classes: 52.6% are not spliced, 45.1% have a large exon with short additional 5' or 3' exons, and 2.3% have a multiple exon structure more typical of higher eukaryotes.

Animals↗

DNA sequencing and analysis of a 67.4 kb region from the right arm of Schizosaccharomyces pombe chromosome II reveals 28 open reading frames including the genes his5, pol5, ppa2, rip1, rpb8 and skb1.

67 393 bp of contiguous DNA located between markers cdc18 and cdc14 on the right arm of fission yeast chromosome II has been sequenced as part of the European Union Schizosaccharomyces pombe genome sequencing project. The complete sequence, contained in cosmid clones c15C4 and c21H7, has been determined on both strands. Sequence analysis shows that it contains 28 open reading frames capable of coding for proteins, 16 split by one or more introns, but no tRNA, rRNA or transposon sequences. The gene density is one per 2. 4 kb. Six genes have been previously described (his5, pol5, ppa2, rip1, rpb8 and skb1) and 22 are novel. Of the novel genes, 14 have significant similarity with proteins of known function, three have similarities with proteins of unknown function and five show no extensive similarities with known proteins. Sequence similarities suggest that three of the novel genes encode ATP-dependent RNA helicases, two encode transcription factor components and others encode a G-protein, a dehydrogenase, a Rab escort protein, an Abc1-like protein, a lipase, an ATP-binding transport protein, an amino acid permease, an acid phosphatase and a mannosyltransferase.

Chromosome Mapping↗

Regulation of hmp gene transcription in Mycobacterium tuberculosis: effects of oxygen limitation and nitrosative and oxidative stress.

The Mycobacterium tuberculosis hmp gene encodes a protein which is homologous to flavohemoglobin in Escherichia coli. Northern blotting analysis demonstrated that hmp transcription increased when a microaerophilic culture became oxygen limited as it entered stationary phase at 20 days. There was a fivefold increase of the hmp transcripts during early stationary phase compared with the value which was observed in the exponential growth phase. This induction of hmp transcription was not due to changes in the mRNA stability since the half-life of hmp mRNA was very short in a 20-day microaerophilic culture. No induction of hmp mRNA was observed during entry into stationary phase when the culture was continuously aerated. hmp transcription was induced after a short exposure of a late-exponential-phase culture to anaerobic conditions. These data indicate that oxygen limitation is the trigger for hmp gene transcription. In addition, when a microaerophilic culture entered into the stationary phase at 20 days, transcription of hmp increased to a small extent after exposure to S-nitrosoglutathione (a nitric oxide [NO] releaser) and sodium nitroprusside (an NO+ donor) and decreased after exposure to paraquat (a superoxide generator) and H2O2. In log phase (4 days) and late stationary phase (40 days), the transcription of hmp was unaffected by nitrosative and oxidative stress. Three primer extension products were observed. The -10 region is 100% identical to that of promoter T3 in mycobacteria and shows a strong similarity to the -10 sequence of hmp and rpoS promoters in E. coli. These observations of hmp mRNA induction in response to O2 limitation and nitrosative stress suggest that the hmp gene of M. tuberculosis may have a role in protection of the organism from NO killing under microaerophilic conditions.

Aerobiosis↗

Deciphering the biology of Mycobacterium tuberculosis from the complete genome sequence.

Countless millions of people have died from tuberculosis, a chronic infectious disease caused by the tubercle bacillus. The complete genome sequence of the best-characterized strain of Mycobacterium tuberculosis, H37Rv, has been determined and analysed in order to improve our understanding of the biology of this slow-growing pathogen and to help the conception of new prophylactic and therapeutic interventions. The genome comprises 4,411,529 base pairs, contains around 4,000 genes, and has a very high guanine + cytosine content that is reflected in the biased amino-acid content of the proteins. M. tuberculosis differs radically from other bacteria in that a very large portion of its coding capacity is devoted to the production of enzymes involved in lipogenesis and lipolysis, and to two new families of glycine-rich proteins with a repetitive structure that may represent a source of antigenic variation.

Chromosome Mapping↗