PubMed Health⌕ Search

Biomedical subjects

João C Setubal

Publications and source records attributed to João C Setubal.

7 recordsLinked to original sources

Lipoprotein computational prediction in spirochaetal genomes.

Lipoproteins are of great interest in understanding the molecular pathogenesis of spirochaetes. Because spirochaete lipobox sequences exhibit more plasticity than those of other bacteria, application of existing prediction algorithms to emerging sequence data has been problematic. In this paper a novel lipoprotein prediction algorithm is described, designated SpLip, constructed as a hybrid of a lipobox weight matrix approach supplemented by a set of lipoprotein signal peptide rules allowing for conservative amino acid substitutions. Both the weight matrix and the rules are based on a training set of 28 experimentally verified spirochaetal lipoproteins. The performance of the SpLip algorithm was compared to that of the hidden Markov model-based LipoP program and the rules-based algorithm Psort for all predicted protein-coding genes of Leptospira interrogans sv. Copenhageni, L. interrogans sv. Lai, Borrelia burgdorferi, Borrelia garinii, Treponema pallidum and Treponema denticola. Psort sensitivity (13-35 %) was considerably less than that of SpLip (93-100 %) or LipoP (50-84 %) due in part to the requirement of Psort for Ala or Gly at the -1 position, a rule based on E. coli lipoproteins. The percentage of false-positive lipoprotein predictions by the LipoP algorithm (8-30 %) was greater than that of SpLip (0-1 %) or Psort (4-27 %), due in part to the lack of rules in LipoP excluding unprecedented amino acids such as Lys and Arg in the -1 position. This analysis revealed a higher number of predicted spirochaetal lipoproteins than was previously known. The improved performance of the SpLip algorithm provides a more accurate prediction of the complete lipoprotein repertoire of spirochaetes. The hybrid approach of supplementing weight matrix scoring with rules based on knowledge of protein secretion biochemistry may be a general strategy for development of improved prediction algorithms.

Algorithms↗

A framework based on Web service orchestration for bioinformatics workflow management.

Bioinformatics activities are growing all over the world, with proliferation of data and tools. This brings new challenges: how to understand and organize these resources and how to provide interoperability among tools to achieve a given goal. We defined and implemented a framework to help meet some of these challenges. Four issues were considered: the use of Web services as a basic unit, the notion of a Semantic Web to improve interoperability at the syntactic and semantic levels, and the use of scientific workflows to coordinate services to be executed, including their interdependencies and service orchestration.

Algorithms↗

Bacterial phytopathogens and genome science.

There are now fourteen completed genomes of bacterial phytopathogens, all of which have been generated in the past six years. These genomes come from a phylogenetically diverse set of organisms, and range in size from 870 kb to more than 6Mb. The publication of these annotated genomes has significantly helped our understanding of bacterial plant disease. These genomes have also provided important information about bacterial evolution. Examples of recently completed genomes include: Pseudomonas syringae pv tomato, which is notable for its large repertoire of effector proteins; Leifsonia xyli subsp. xyli, the first Gram-positive bacterial genome to be sequenced; and Phytoplasma asteris, the small genome that lacks important functions previously thought to be essential in a bacterium.

Actinomycetales↗

Comparative analyses of Xanthomonas and Xylella complete genomes.

Computational analyses of four bacterial genomes of the Xanthomonadaceae family reveal new unique genes that may be involved in adaptation, pathogenicity, and host specificity. The Xanthomonas genus presents 3636 unique genes distributed in 1470 families, while Xylella genus presents 1026 unique genes distributed in 375 families. Among Xanthomonas-specific genes, we highlight a large number of cell wall degrading enzymes, proteases, and iron receptors, a set of energy metabolism genes, second copy of the type II secretion system, type III secretion system, flagella and chemotactic machinery, and the xanthomonadin synthesis gene cluster. Important genes unique to the Xylella genus are an additional copy of a type IV pili gene cluster and the complete machinery of colicin V synthesis and secretion. Intersections of gene sets from both genera reveal a cluster of genes homologous to Salmonella's SPI-7 island in Xanthomonas axonopodis pv citri and Xylella fastidiosa 9a5c, which might be involved in host specificity. Each genome also presents important unique genes, such as an HMS cluster, the kdgT gene, and O-antigen in Xanthomonas axonopodis pv citri; a number of avrBS genes and a distinct O-antigen in Xanthomonas campestris pv campestris, a type I restriction-modification system and a nickase gene in Xylella fastidiosa 9a5c, and a type II restriction-modification system and two genes related to peptidoglycan biosynthesis in Xylella fastidiosa temecula 1. All these differences imply a considerable number of gene gains and losses during the divergence of the four lineages, and are associated with structural genome modifications that may have a direct relation with the mode of transmission, adaptation to specific environments and pathogenicity of each organism.

Bacterial Physiological Phenomena↗

Saci-1, -2, and -3 and Perere, four novel retrotransposons with high transcriptional activities from the human parasite Schistosoma mansoni.

Using the data set of 180,000 expressed sequence tags (ESTs) of the blood fluke Schistosoma mansoni generated recently by our group, we identified three novel long-terminal-repeat (LTR)- and one novel non-LTR-expressed retrotransposon, named Saci-1, -2, and -3 and Perere, respectively. Full-length sequences were reconstructed from ESTs and have deduced open reading frames (ORFs) with several uncorrupted features, characterizing them as possible active retrotransposons of different known transposon families. Alignment of reconstructed sequences to available preliminary genome sequence data confirmed the overall structure of the transposons. The frequency of sequenced transposon transcripts in cercariae was 14% of all transcripts from that stage, twofold higher than that in schistosomula and three- to fourfold higher than that in adults, eggs, miracidia, and germ balls. We show by Southern blot analysis, by EST annotation and tallying, and by counting transposon tags from a Serial Analysis of Gene Expression library, that the four novel retrotransposons exhibit a 10- to 30-fold lower copy number in the genome and a 4- to 200-fold-higher transcriptional rate per copy than the four previously described S. mansoni retrotransposons [corrected]. Such differences lead us to hypothesize that there are two different populations of retrotransposons in S. mansoni genome, occupying different niches in its ecology. Examples of retrotransposon fragment inserts were found into the 5' and 3' untranslated regions of four different S. mansoni target gene transcripts. The data presented here suggest a role for these elements in the dynamics of this complex human parasite genome.

Amino Acid Sequence↗

Comparative genomics analyses of citrus-associated bacteria.

Xylella fastidiosa 9a5c (XF-9a5c) and Xanthomonas axonopodis pv. citri (XAC) are bacteria that infect citrus plants. Sequencing of the genomes of these strains is complete and comparative analyses are now under way with the genomes of other bacteria of the same genera. In this review, we present an overview of this comparative genomic work. We also present a detailed genomic comparison between XF-9a5a and XAC. Based on this analysis, genes and operons were identified that might be relevant for adaptation to citrus. XAC has two copies of a type II secretion system, a large number of cell wall-degrading enzymes and sugar transporters, a complete energy metabolism, a whole set of avirulence genes associated with a type III secretion system, and a complete flagellar and chemotatic system. By contrast, XF-9a5c possesses more genes involved with type IV pili biosynthesis than does XAC, contains genes encoding for production of colicins, and has 4 copies of Type I restriction/modification system while XAC has only one.

Animals↗

Transcriptome analysis of the acoelomate human parasite Schistosoma mansoni.

Schistosoma mansoni is the primary causative agent of schistosomiasis, which affects 200 million individuals in 74 countries. We generated 163,000 expressed-sequence tags (ESTs) from normalized cDNA libraries from six selected developmental stages of the parasite, resulting in 31,000 assembled sequences and 92% sampling of an estimated 14,000 gene complement. By analyzing automated Gene Ontology assignments, we provide a detailed view of important S. mansoni biological systems, including characterization of metazoa-specific and eukarya-conserved genes. Phylogenetic analysis suggests an early divergence from other metazoa. The data set provides insights into the molecular mechanisms of tissue organization, development, signaling, sexual dimorphism, host interactions and immune evasion and identifies novel proteins to be investigated as vaccine candidates and potential drug targets.

Animals↗