PubMed Health⌕ Search

Biomedical subjects

N Mulder

Publications and source records attributed to N Mulder.

7 recordsLinked to original sources

The transcriptional landscape of the mammalian genome.

This study describes comprehensive polling of transcription start and termination sites and analysis of previously unidentified full-length complementary DNAs derived from the mouse genome. We identify the 5' and 3' boundaries of 181,047 transcripts with extensive variation in transcripts arising from alternative promoter usage, splicing, and polyadenylation. There are 16,247 new mouse protein-coding transcripts, including 5154 encoding previously unidentified proteins. Genomic mapping of the transcriptome reveals transcriptional forests, with overlapping transcription on both strands, separated by deserts in which few transcripts are observed. The data provide a comprehensive platform for the comparative analysis of mammalian transcriptional regulation in differentiation and development.

3' Untranslated Regions↗

InterProScan: protein domains identifier.

InterProScan [E. M. Zdobnov and R. Apweiler (2001) Bioinformatics, 17, 847-848] is a tool that combines different protein signature recognition methods from the InterPro [N. J. Mulder, R. Apweiler, T. K. Attwood, A. Bairoch, A. Bateman, D. Binns, P. Bradley, P. Bork, P. Bucher, L. Cerutti et al. (2005) Nucleic Acids Res., 33, D201-D205] consortium member databases into one resource. At the time of writing there are 10 distinct publicly available databases in the application. Protein as well as DNA sequences can be analysed. A web-based version is accessible for academic and commercial organizations from the EBI (http://www.ebi.ac.uk/InterProScan/). In addition, a standalone Perl version and a SOAP Web Service [J. Snell, D. Tidwell and P. Kulchenko (2001) Programming Web Services with SOAP, 1st edn. O'Reilly Publishers, Sebastopol, CA, http://www.w3.org/TR/soap/] are also available to the users. Various output formats are supported and include text tables, XML documents, as well as various graphs to help interpret the results.

Databases, Protein↗

Interactive InterPro-based comparisons of proteins in whole genomes.

MOTIVATION: The SWISS-PROT group at the EBI has developed the Proteome Analysis Database utilizing existing resources and providing comprehensive and integrated comparative analysis of the predicted protein coding sequences of the complete genomes of bacteria, archaea and eukaryotes. The Proteome Analysis Database is accompanied by a program that has been designed to carry out interactive InterPro proteome comparisons for any one proteome against any other one or more of the proteomes in the database.

Computational Biology↗

Proteome Analysis Database: online application of InterPro and CluSTr for the functional classification of proteins in whole genomes.

The SWISS-PROT group at EBI has developed the Proteome Analysis Database utilising existing resources and providing comparative analysis of the predicted protein coding sequences of the complete genomes of bacteria, archaea and eukaryotes (http://www.ebi.ac. uk/proteome/). The two main projects used, InterPro and CluSTr, give a new perspective on families, domains and sites and cover 31-67% (InterPro statistics) of the proteins from each of the complete genomes. CluSTr covers the three complete eukaryotic genomes and the incomplete human genome data. The Proteome Analysis Database is accompanied by a program that has been designed to carry out InterPro proteome comparisons for any one proteome against any other one or more of the proteomes in the database.

Animals↗

Profiling the malaria genome: a gene survey of three species of malaria parasite with comparison to other apicomplexan species.

We have undertaken the first comparative pilot gene discovery analysis of approximately 25,000 random genomic and expressed sequence tags (ESTs) from three species of Plasmodium, the infectious agent that causes malaria. A total of 5482 genome survey sequences (GSSs) and 5582 ESTs were generated from mung bean nuclease (MBN) and cDNA libraries, respectively, of the ANKA line of the rodent malaria parasite Plasmodium berghei, and 10,874 GSSs generated from MBN libraries of the Salvador I and Belem lines of Plasmodium vivax, the most geographically wide-spread human malaria pathogen. These tags, together with 2438 Plasmodium falciparum sequences present in GenBank, were used to perform first-pass assembly and transcript reconstruction, and non-redundant consensus sequence datasets created. The datasets were compared against public protein databases and more than 1000 putative new Plasmodium proteins identified based on sequence similarity. Homologs of previously characterized Plasmodium genes were also identified, increasing the number of P. vivax and P. berghei sequences in public databases at least 10-fold. Comparative studies with other species of Apicomplexa identified interesting homologs of possible therapeutic or diagnostic value. A gene prediction program, Phat, was used to predict probable open reading frames for proteins in all three datasets. Predicted and non-redundant BLAST-matched proteins were submitted to InterPro, an integrated database of protein domains, signatures and families, for functional classification. Thus a partial predicted proteome was created for each species. This first comparative analysis of Plasmodium protein coding sequences represents a valuable resource for further studies on the biology of this important pathogen.

Animals↗

An EORTC Gastrointestinal Group evaluation of the combination of sequential methotrexate and 5-fluorouracil, combined with adriamycin in advanced measurable gastric cancer.

In a phase II multicenter trial, 71 patients with advanced measurable gastric cancer were registered to receive sequential high-dose methotrexate (MTX) and 5-fluorouracil (5-FU) combined with Adriamycin (A [Adria Laboratories, Columbus, OH]). The response rate was 33% (22 of 67), including all eligible patients. There were nine complete responders (CRs). The median survival for all patients was 6 months. There has been one toxic death; however, three other patients died from toxicity associated with major protocol violations. It is concluded that this protocol is active in gastric cancer. Toxicity, partly because of nonprotocol adherence, is considerable and is now under further investigation in a randomized trial comparing this schedule with a combination of 5-FU, Adriamycin, and mitomycin C (FAM).

Adult↗

[InterPro as a new tool for whole genome analysis. A comparative analysis of Mycobacterium tuberculosis, Bacillus subtilis and Escherichia coli as a case study].

InterPro was developed as a new integrated documentation resource for protein families, domains and functional sites to rationalize the complementary efforts of the PROSITE, PRINTS, Pfam and ProDom database projects and has applications in computational functional classification of newly determined sequences lacking biochemical characterization and in comparative genome analysis. InterPro contains over 3500 entries, with more than 1000000 hits in SWISS-PROT and TrEMBL. The database is accessible for text- and sequence-based searches at http://www.ebi.ac.uk/interpro/. InterPro was used for whole proteome analysis of the pathogenic microorganism, Mycobacterium tuberculosis, and comparison with the predicted protein coding sequences of the complete genomes of Bacillus subtilis and Escherichia coli. 64.8% of the M. tuberculosis proteins in the proteome matched InterPro entries, and these could be classified according to function. The comparison with B. subtilis and E. coli provided information on the most common protein families and domains, and the most highly represented families in each organism. InterPro thus provides a useful tool for global views of whole proteomes and their compositions.

Bacillus subtilis↗