PubMed Health⌕ Search

Biomedical subjects

Daniela Bartels

Publications and source records attributed to Daniela Bartels.

12 recordsLinked to original sources

Whole-genome sequence of Listeria welshimeri reveals common steps in genome reduction with Listeria innocua as compared to Listeria monocytogenes.

We present the complete genome sequence of Listeria welshimeri, a nonpathogenic member of the genus Listeria. Listeria welshimeri harbors a circular chromosome of 2,814,130 bp with 2,780 open reading frames. Comparative genomic analysis of chromosomal regions between L. welshimeri, Listeria innocua, and Listeria monocytogenes shows strong overall conservation of synteny, with the exception of the translocation of an F(o)F(1) ATP synthase. The smaller size of the L. welshimeri genome is the result of deletions in all of the genes involved in virulence and of "fitness" genes required for intracellular survival, transcription factors, and LPXTG- and LRR-containing proteins as well as 55 genes involved in carbohydrate transport and metabolism. In total, 482 genes are absent from L. welshimeri relative to L. monocytogenes. Of these, 249 deletions are commonly absent in both L. welshimeri and L. innocua, suggesting similar genome evolutionary paths from an ancestor. We also identified 311 genes specific to L. welshimeri that are absent in the other two species, indicating gene expansion in L. welshimeri, including horizontal gene transfer. The species L. welshimeri appears to have been derived from early evolutionary events and an ancestor more compact than L. monocytogenes that led to the emergence of nonpathogenic Listeria spp.

Chromosomes, Bacterial↗

Genome sequence of the ubiquitous hydrocarbon-degrading marine bacterium Alcanivorax borkumensis.

Alcanivorax borkumensis is a cosmopolitan marine bacterium that uses oil hydrocarbons as its exclusive source of carbon and energy. Although barely detectable in unpolluted environments, A. borkumensis becomes the dominant microbe in oil-polluted waters. A. borkumensis SK2 has a streamlined genome with a paucity of mobile genetic elements and energy generation-related genes, but with a plethora of genes accounting for its wide hydrocarbon substrate range and efficient oil-degradation capabilities. The genome further specifies systems for scavenging of nutrients, particularly organic and inorganic nitrogen and oligo-elements, biofilm formation at the oil-water interface, biosurfactant production and niche-specific stress responses. The unique combination of these features provides A. borkumensis SK2 with a competitive edge in oil-polluted environments. This genome sequence provides the basis for the future design of strategies to mitigate the ecological damage caused by oil spills.

Base Sequence↗

Finding novel genes in bacterial communities isolated from the environment.

MOTIVATION: Novel sequencing techniques can give access to organisms that are difficult to cultivate using conventional methods. When applied to environmental samples, the data generated has some drawbacks, e.g. short length of assembled contigs, in-frame stop codons and frame shifts. Unfortunately, current gene finders cannot circumvent these difficulties. At the same time, the automated prediction of genes is a prerequisite for the increasing amount of genomic sequences to ensure progress in metagenomics. RESULTS: We introduce a novel gene finding algorithm that incorporates features overcoming the short length of the assembled contigs from environmental data, in-frame stop codons as well as frame shifts contained in bacterial sequences. The results show that by searching for sequence similarities in an environmental sample our algorithm is capable of detecting a high fraction of its gene content, depending on the species composition and the overall size of the sample. The method is valuable for hunting novel unknown genes that may be specific for the habitat where the sample is taken. Finally, we show that our algorithm can even exploit the limited information contained in the short reads generated by 454 technology for the prediction of protein coding genes. AVAILABILITY: The program is freely available upon request.

Algorithms↗

Physical map of the Azoarcus sp. strain BH72 genome based on a bacterial artificial chromosome library as a platform for genome sequencing and functional analysis.

Azoarcus sp. strain BH72 is a Gram-negative proteobacterium of the beta subclass; it is a diazotrophic endophyte of graminaceous plants and can provide significant amounts of fixed nitrogen to its host plant Kallar grass. We aimed to obtain a physical map of the Azoarcus sp. strain BH72 chromosome to be directly used in functional analysis and as a part of an Azoarcus sp. BH72 genome project. A bacterial artificial chromosome (BAC) library was constructed and analysed. A representative physical map with a high density of marker genes was developed in which 64 aligned BAC clones covered almost the entire genome.

Azoarcus↗

Insights into genome plasticity and pathogenicity of the plant pathogenic bacterium Xanthomonas campestris pv. vesicatoria revealed by the complete genome sequence.

The gram-negative plant-pathogenic bacterium Xanthomonas campestris pv. vesicatoria is the causative agent of bacterial spot disease in pepper and tomato plants, which leads to economically important yield losses. This pathosystem has become a well-established model for studying bacterial infection strategies. Here, we present the whole-genome sequence of the pepper-pathogenic Xanthomonas campestris pv. vesicatoria strain 85-10, which comprises a 5.17-Mb circular chromosome and four plasmids. The genome has a high G+C content (64.75%) and signatures of extensive genome plasticity. Whole-genome comparisons revealed a gene order similar to both Xanthomonas axonopodis pv. citri and Xanthomonas campestris pv. campestris and a structure completely different from Xanthomonas oryzae pv. oryzae. A total of 548 coding sequences (12.2%) are unique to X. campestris pv. vesicatoria. In addition to a type III secretion system, which is essential for pathogenicity, the genome of strain 85-10 encodes all other types of protein secretion systems described so far in gram-negative bacteria. Remarkably, one of the putative type IV secretion systems encoded on the largest plasmid is similar to the Icm/Dot systems of the human pathogens Legionella pneumophila and Coxiella burnetii. Comparisons with other completely sequenced plant pathogens predicted six novel type III effector proteins and several other virulence factors, including adhesins, cell wall-degrading enzymes, and extracellular polysaccharides.

Adhesins, Bacterial↗

BACCardI--a tool for the validation of genomic assemblies, assisting genome finishing and intergenome comparison.

SUMMARY: We provide the graphical tool BACCardI for the construction of virtual clone maps from standard assembler output files or BLAST based sequence comparisons. This new tool has been applied to numerous genome projects to solve various problems including (a) validation of whole genome shotgun assemblies, (b) support for contig ordering in the finishing phase of a genome project, and (c) intergenome comparison between related strains when only one of the strains has been sequenced and a large insert library is available for the other. The BACCardI software can seamlessly interact with various sequence assembly packages. MOTIVATION: Genomic assemblies generated from sequence information need to be validated by independent methods such as physical maps. The time-consuming task of building physical maps can be circumvented by virtual clone maps derived from read pair information of large insert libraries.

Algorithms↗

Building a BRIDGE for the integration of heterogeneous data from functional genomics into a platform for systems biology.

The flood of data acquired from the increasing number of publicly available genomes has led to new demands for bioinformatics software. With the growing amount of information resulting from high throughput experiments new questions arise that often focus on the comparison of genes, genomes, and their expression profiles. Inferring new knowledge by combining different kinds of "post-genomics" data obviously necessitates the development of new approaches that allow the integration of variable data sources into a flexible framework. In this paper, we describe our concept for the integration of heterogeneous data into a platform for systems biology. We have implemented a Bioinformatics Resource for the Integration of heterogeneous Data from Genomic Explorations (BRIDGE) and illustrate the usability of our approach as a platform for systems biology for two sample applications.

Algorithms↗

Whole genome shotgun sequencing guided by bioinformatics pipelines--an optimized approach for an established technique.

While the sequencing of bacterial genomes has become a routine procedure at major sequencing centers, there are still a number of genome projects at small- or medium-size facilities. For these facilities a maximum of control over sequencing, assembling and finishing is essential. At the same time, facilities have to be able to co-operate at minimum costs for the overall project. We have established a pipeline for the distributed sequencing of Alcanivorax borkumensis SK2, Azoarcus sp. BH72, Clavibacter michiganensis subsp. michiganensis NCPPB382, Sorangium cellulosum So ce56 and Xanthomonas campestris pv. vesicatoria 85-10. Our pipeline relies on standard tools (e.g. PHRED/PHRAP, CAP3 and Consed/Autofinish) wherever possible, supplementing them with new tools (BioMake and BACCardI) to achieve the aims described above.

Algorithms↗

Bioinformatics support for high-throughput proteomics.

In the "post-genome" era, mass spectrometry (MS) has become an important method for the analysis of proteome data. The rapid advancement of this technique in combination with other methods used in proteomics results in an increasing number of high-throughput projects. This leads to an increasing amount of data that needs to be archived and analyzed. To cope with the need for automated data conversion, storage, and analysis in the field of proteomics, the open source system ProDB was developed. The system handles data conversion from different mass spectrometer software, automates data analysis, and allows the annotation of MS spectra (e.g. assign gene names, store data on protein modifications). The system is based on an extensible relational database to store the mass spectra together with the experimental setup. It also provides a graphical user interface (GUI) for managing the experimental steps which led to the MS data. Furthermore, it allows the integration of genome and proteome data. Data from an ongoing experiment was used to compare manual and automated analysis. First tests showed that the automation resulted in a significant saving of time. Furthermore, the quality and interpretability of the results was improved in all cases.

Algorithms↗

EMMA: a platform for consistent storage and efficient analysis of microarray data.

As a high throughput technique, microarray experiments produce large data sets, consisting of measured data, laboratory protocols, and experimental settings. We have implemented the open source platform EMMA to store and analyze these data. The system provides automated pipelines for data processing and has a modular architecture that can be easily extended. EMMA features detailed reports about spots and their corresponding measurements. In addition to routine data analysis algorithms, the system can be integrated with other components that contain additional data sources (e.g. genome annotation systems).

Algorithms↗

The complete Corynebacterium glutamicum ATCC 13032 genome sequence and its impact on the production of L-aspartate-derived amino acids and vitamins.

The complete genomic sequence of Corynebacterium glutamicum ATCC 13032, well-known in industry for the production of amino acids, e.g. of L-glutamate and L-lysine was determined. The C. glutamicum genome was found to consist of a single circular chromosome comprising 3282708 base pairs. Several DNA regions of unusual composition were identified that were potentially acquired by horizontal gene transfer, e.g. a segment of DNA from C. diphtheriae and a prophage-containing region. After automated and manual annotation, 3002 protein-coding genes have been identified, and to 2489 of these, functions were assigned by homologies to known proteins. These analyses confirm the taxonomic position of C. glutamicum as related to Mycobacteria and show a broad metabolic diversity as expected for a bacterium living in the soil. As an example for biotechnological application the complete genome sequence was used to reconstruct the metabolic flow of carbon into a number of industrially important products derived from the amino acid L-aspartate.

Amino Acid Sequence↗

GenDB--an open source genome annotation system for prokaryote genomes.

The flood of sequence data resulting from the large number of current genome projects has increased the need for a flexible, open source genome annotation system, which so far has not existed. To account for the individual needs of different projects, such a system should be modular and easily extensible. We present a genome annotation system for prokaryote genomes, which is well tested and readily adaptable to different tasks. The modular system was developed using an object-oriented approach, and it relies on a relational database backend. Using a well defined application programmers interface (API), the system can be linked easily to other systems. GenDB supports manual as well as automatic annotation strategies. The software currently is in use in more than a dozen microbial genome annotation projects. In addition to its use as a production genome annotation system, it can be employed as a flexible framework for the large-scale evaluation of different annotation strategies. The system is open source.

Amino Acid Sequence↗