PubMed Health⌕ Search

Biomedical subjects

Emmanuel Mongin

Publications and source records attributed to Emmanuel Mongin.

9 recordsLinked to original sources

BASC: an integrated bioinformatics system for Brassica research.

The BASC system provides tools for the integrated mining and browsing of genetic, genomic and phenotypic data. This public resource hosts information on Brassica species supporting the Multinational Brassica Genome Sequencing Project, and is based upon five distinct modules, ESTDB, Microarray, MarkerQTL, CMap and EnsEMBL. ESTDB hosts expressed gene sequences and related annotation derived from comparison with GenBank, UniRef and the genome sequence of Arabidopsis. The Microarray module hosts gene expression information related to genes annotated within ESTDB. MarkerQTL is the most complex module and integrates information on genetic markers, maps, individuals, genotypes and traits. Two further modules include an Arabidopsis EnsEMBL genome viewer and the CMap comparative genetic map viewer for the visualization and integration of genetic and genomic data. The database is accessible at http://bioinformatics.pbcbasc.latrobe.edu.au.

Arabidopsis↗

Anopheles gambiae immune responses to human and rodent Plasmodium parasite species.

Transmission of malaria is dependent on the successful completion of the Plasmodium lifecycle in the Anopheles vector. Major obstacles are encountered in the midgut tissue, where most parasites are killed by the mosquito's immune system. In the present study, DNA microarray analyses have been used to compare Anopheles gambiae responses to invasion of the midgut epithelium by the ookinete stage of the human pathogen Plasmodium falciparum and the rodent experimental model pathogen P. berghei. Invasion by P. berghei had a more profound impact on the mosquito transcriptome, including a variety of functional gene classes, while P. falciparum elicited a broader immune response at the gene transcript level. Ingestion of human malaria-infected blood lacking invasive ookinetes also induced a variety of immune genes, including several anti-Plasmodium factors. Twelve selected genes were assessed for effect on infection with both parasite species and bacteria using RNAi gene silencing assays, and seven of these genes were found to influence mosquito resistance to both parasite species. An MD2-like receptor, AgMDL1, and an immunolectin, FBN39, showed specificity in regulating only resistance to P. falciparum, while the antimicrobial peptide gambicin and a novel putative short secreted peptide, IRSP5, were more specific for defense against the rodent parasite P. berghei. While all the genes that affected Plasmodium development also influenced mosquito resistance to bacterial infection, four of the antimicrobial genes had no effect on Plasmodium development. Our study shows that the impact of P. falciparum and P. berghei infection on A. gambiae biology at the gene transcript level is quite diverse, and the defense against the two Plasmodium species is mediated by antimicrobial factors with both universal and Plasmodium-species specific activities. Furthermore, our data indicate that the mosquito is capable of sensing infected blood constituents in the absence of invading ookinetes, thereby inducing anti-Plasmodium immune responses.

Animals↗

SNPServer: a real-time SNP discovery tool.

SNPServer is a real-time flexible tool for the discovery of SNPs (single nucleotide polymorphisms) within DNA sequence data. The program uses BLAST, to identify related sequences, and CAP3, to cluster and align these sequences. The alignments are parsed to the SNP discovery software autoSNP, a program that detects SNPs and insertion/deletion polymorphisms (indels). Alternatively, lists of related sequences or pre-assembled sequences may be entered for SNP discovery. SNPServer and autoSNP use redundancy to differentiate between candidate SNPs and sequence errors. For each candidate SNP, two measures of confidence are calculated, the redundancy of the polymorphism at a SNP locus and the co-segregation of the candidate SNP with other SNPs in the alignment. SNPServer is available at http://hornbill.cspp.latrobe.edu.au/snpdiscovery.html.

Internet↗

An overview of Ensembl.

Ensembl (http://www.ensembl.org/) is a bioinformatics project to organize biological information around the sequences of large genomes. It is a comprehensive source of stable automatic annotation of individual genomes, and of the synteny and orthology relationships between them. It is also a framework for integration of any biological data that can be mapped onto features derived from the genomic sequence. Ensembl is available as an interactive Web site, a set of flat files, and as a complete, portable open source software system for handling genomes. All data are provided without restriction, and code is freely available. Ensembl's aims are to continue to "widen" this biological integration to include other model organisms relevant to understanding human biology as they become available; to "deepen" this integration to provide an ever more seamless linkage between equivalent components in different species; and to provide further classification of functional elements in the genome that have been previously elusive.

Computational Biology↗

The Anopheles gambiae genome: an update.

As a result of an international collaborative effort, the first draft of the Anopheles gambiae genome sequence and its preliminary annotation were published in October 2002. Since then, the assembly, annotation and means of accession of the An. gambiae genome have been under continuous development. This article reviews progress and considers limitations in the current sequence assembly and gene annotation, as well as approaches to address these problems and outstanding issues that users of the data must bear in mind.

Animals↗

The Ensembl automatic gene annotation system.

As more genomes are sequenced, there is an increasing need for automated first-pass annotation which allows timely access to important genomic information. The Ensembl gene-building system enables fast automated annotation of eukaryotic genomes. It annotates genes based on evidence derived from known protein, cDNA, and EST sequences. The gene-building system rests on top of the core Ensembl (MySQL) database schema and Perl Application Programming Interface (API), and the data generated are accessible through the Ensembl genome browser (http://www.ensembl.org). To date, the Ensembl predicted gene sets are available for the A. gambiae, C. briggsae, zebrafish, mouse, rat, and human genomes and have been heavily relied upon in the publication of the human, mouse, rat, and A. gambiae genome sequence analysis. Here we describe in detail the gene-building system and the algorithms involved. All code and data are freely available from http://www.ensembl.org.

Animals↗

The Ensembl analysis pipeline.

The Ensembl pipeline is an extension to the Ensembl system which allows automated annotation of genomic sequence. The software comprises two parts. First, there is a set of Perl modules ("Runnables" and "RunnableDBs") which are 'wrappers' for a variety of commonly used analysis tools. These retrieve sequence data from a relational database, run the analysis, and write the results back to the database. They inherit from a common interface, which simplifies the writing of new wrapper modules. On top of this sits a job submission system (the "RuleManager") which allows efficient and reliable submission of large numbers of jobs to a compute farm. Here we describe the fundamental software components of the pipeline, and we also highlight some features of the Sanger installation which were necessary to enable the pipeline to scale to whole-genome analysis.

Base Sequence↗

The genome sequence of the malaria mosquito Anopheles gambiae.

Anopheles gambiae is the principal vector of malaria, a disease that afflicts more than 500 million people and causes more than 1 million deaths each year. Tenfold shotgun sequence coverage was obtained from the PEST strain of A. gambiae and assembled into scaffolds that span 278 million base pairs. A total of 91% of the genome was organized in 303 scaffolds; the largest scaffold was 23.1 million base pairs. There was substantial genetic variation within this strain, and the apparent existence of two haplotypes of approximately equal frequency ("dual haplotypes") in a substantial fraction of the genome likely reflects the outbred nature of the PEST strain. The sequence produced a conservative inference of more than 400,000 single-nucleotide polymorphisms that showed a markedly bimodal density distribution. Analysis of the genome sequence revealed strong evidence for about 14,000 protein-encoding transcripts. Prominent expansions in specific families of proteins likely involved in cell adhesion and immunity were noted. An expressed sequence tag analysis of genes regulated by blood feeding provided insights into the physiological adaptations of a hematophagous insect.

Animals↗

Genome annotation techniques: new approaches and challenges.

As more of the human genome draft sequence is finished, and genomes from other organisms begin to be sequenced, the demand for accurate and reliable genome annotation will increase significantly. To facilitate this industrial-scale genome annotation, automated bioinformatics solutions are increasingly required. As a result, automatic genome annotation systems have become more important in gene discovery within recent years. The design of such large-scale bioinformatics systems is an evolving and dynamic field, based on central cores of bioinformatics software tools and relational databases. Not only must these systems efficiently manage and integrate large volumes of genomic data, but they must also deliver accurate gene predictions and effectively distribute annotation data to the biosciences community.

Computational Biology↗